Clipping-Based Post Training 8-Bit Quantization of Convolution Neural Networks for Object Detection

Leisheng Chen,Peihuang Lou

doi:10.3390/app122312405

Leisheng Chen, Peihuang Lou

Open Access

https://doi.org/10.3390/app122312405

Copy DOI

Abstract

Fueled by the development of deep neural networks, breakthroughs have been achieved in plenty of computer vision problems, such as image classification, segmentation, and object detection. These models usually have handers and millions of parameters, which makes them both computational and memory expensive. Motivated by this, this paper proposes a post-training quantization method based on the clipping operation for neural network compression. By quantizing parameters of a model to 8-bit using our proposed methods, its memory consumption is reduced, its computational speed is increased, and its performance is maintained. This method exploits the clipping operation during training so that it saves a large computational cost during quantization. After training, this method quantizes the parameters to 8-bit based on the clipping value. In addition, a fully connected layer compression is conducted using singular value decomposition (SVD), and a novel loss function term is leveraged to further diminish the performance drop caused by quantization. The proposed method is validated on two widely used models, Yolo V3 and Faster R-CNN, for object detection on the PASCAL VOC, COCO, and ImageNet datasets. Performances show it effectively reduces the storage consumption at 18.84% and accelerates the model at 381%, meanwhile avoiding the performance drop (drop < 0.02% in VOC).

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

Journal: Applied Sciences	Publication Date: Dec 4, 2022
Citations: 2	License type: CC BY 4.0

R Discovery Prime

R Discovery Prime

Clipping-Based Post Training 8-Bit Quantization of Convolution Neural Networks for Object Detection

Abstract

Talk to us

Similar Papers

More From: Applied Sciences

Lead the way for us

Similar Papers

BDNN: Binary convolution neural networks for fast object detection
Hanyu Peng ... Shifeng Chen
Pattern Recognition Letters | VOL. 125
Hanyu Peng, et. al.Hanyu Peng ... Shifeng Chen
08 Apr 2019
Pattern Recognition Letters | VOL. 125

Quantization and Training of Low Bit-Width Convolutional Neural Networks for Object Detection
Penghang Yin
Journal of Computational Mathematics | VOL. 37
Penghang YinPenghang Yin
01 Jun 2019
Journal of Computational Mathematics | VOL. 37

MRMNet: Multi-scale residual multi-branch neural network for object detection
Yongsheng Dong ... Xuelong Li
Neurocomputing | VOL. 596
Yongsheng Dong, et. al.Yongsheng Dong ... Xuelong Li
01 May 2024
Neurocomputing | VOL. 596

Physically realizable adversarial examples for convolutional object detection algorithms
David R Chambers ... Harold A Garza
-
David R Chambers, et. al.David R Chambers ... Harold A Garza
14 May 2019
14 May 2019

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Clipping-Based Post Training 8-Bit Quantization of Convolution Neural Networks for Object Detection

Abstract

Talk to us

Similar Papers

More From: Applied Sciences