Modified Convolution Neural Network for Highly Effective Parallel Processing

Sang-Soo Park,Ki-Seok Chung,Jung-Hyun Hong

doi:10.1109/iri.2017.37

Abstract

Today, Convolutional Neural Network (CNN) is adopted in a lot of areas such as computer vision and natural language processing. By employing hardware accelerators such as graphic processing unit (GPU), a significant amount of speedup can be achieved in CNN and many studies have proposed such acceleration methods. However, it is not straightforward to parallelize the CNN on a hardware accelerator because there are irregular characteristics of generating output feature maps. In this paper, we propose a modified CNN for efficient parallel processing. A well-known CNN architecture called Lenet-5 has an inefficient convolution combination. The proposed method of this paper improves the efficiency by utilizing a special operation called dummy operation. The proposed method is capable of maximizing the utilization of GPU by modifying Lenet-5's convolution combination. Its improved efficiency is validated on a platform that integrates a CPU and a GPU in the same die. Our OpenCL implementation of the proposed method has achieved an average peak performance of 115.66 GFLOPS which is an improvement of 37.26 times in execution time. Further, a reduction of 26.40 times in energy consumption is achieved.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Modified Convolution Neural Network for Highly Effective Parallel Processing

Abstract

Talk to us

Similar Papers

Lead the way for us

Similar Papers

Accelerating Convolutional Neural Network With FFT on Embedded Hardware
Tahmid Abtahi ... Tinoosh Mohsenin
IEEE Transactions on Very Large Scale Integration (VLSI) Systems | VOL. 26
Tahmid Abtahi, et. al.Tahmid Abtahi ... Tinoosh Mohsenin
01 Sep 2018
IEEE Transactions on Very Large Scale Integration (VLSI) Systems | VOL. 26

Efficient SIMD implementation for accelerating convolutional neural network
Sung-Jin Lee ... Ki-Seok Chung
-
Sung-Jin Lee, et. al.Sung-Jin Lee ... Ki-Seok Chung
02 Nov 2018
02 Nov 2018

Interview
-
Electronics Letters | VOL. 54
--
01 May 2018
Electronics Letters | VOL. 54

RETRACTED ARTICLE: A novel cognitive Wallace compressor based multi operand adders in CNN architecture for FPGA
T Kowsalya
Journal of Ambient Intelligence and Humanized Computing | VOL. 12
T KowsalyaT Kowsalya
07 Aug 2020
Journal of Ambient Intelligence and Humanized Computing | VOL. 12

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Modified Convolution Neural Network for Highly Effective Parallel Processing

Abstract

Talk to us

Similar Papers