Clipping-Based Post Training 8-Bit Quantization of Convolution Neural Networks for Object Detection

被引:2
|
作者
Chen, Leisheng [1 ]
Lou, Peihuang [1 ]
机构
[1] Nanjing Univ Aeronaut & Astronaut, Coll Mech & Elect Engn, Nanjing 210016, Peoples R China
来源
APPLIED SCIENCES-BASEL | 2022年 / 12卷 / 23期
关键词
object detection; quantization; clipping; post-training quantization; accuracy loss;
D O I
10.3390/app122312405
中图分类号
O6 [化学];
学科分类号
0703 ;
摘要
Fueled by the development of deep neural networks, breakthroughs have been achieved in plenty of computer vision problems, such as image classification, segmentation, and object detection. These models usually have handers and millions of parameters, which makes them both computational and memory expensive. Motivated by this, this paper proposes a post-training quantization method based on the clipping operation for neural network compression. By quantizing parameters of a model to 8-bit using our proposed methods, its memory consumption is reduced, its computational speed is increased, and its performance is maintained. This method exploits the clipping operation during training so that it saves a large computational cost during quantization. After training, this method quantizes the parameters to 8-bit based on the clipping value. In addition, a fully connected layer compression is conducted using singular value decomposition (SVD), and a novel loss function term is leveraged to further diminish the performance drop caused by quantization. The proposed method is validated on two widely used models, Yolo V3 and Faster R-CNN, for object detection on the PASCAL VOC, COCO, and ImageNet datasets. Performances show it effectively reduces the storage consumption at 18.84% and accelerates the model at 381%, meanwhile avoiding the performance drop (drop < 0.02% in VOC).
引用
收藏
页数:18
相关论文
共 50 条
  • [41] Edge-Aware Convolution Neural Network Based Salient Object Detection
    Guan, Wenlong
    Wang, Tiantian
    Qi, Jinqing
    Zhang, Lihe
    Lu, Huchuan
    IEEE SIGNAL PROCESSING LETTERS, 2019, 26 (01) : 114 - 118
  • [42] MSQuant: Efficient Post-Training Quantization for Object Detection via Migration Scale Search
    Jiang, Zhesheng
    Li, Chao
    Qu, Tao
    He, Chu
    Wang, Dingwen
    ELECTRONICS, 2025, 14 (03):
  • [43] AVRNTRU: Lightweight NTRU-based Post-Quantum Cryptography for 8-bit AVR Microcontrollers
    Cheng, Iiao
    Grossschadl, Johann
    Ronne, Peter B.
    Ryan, Peter Y. A.
    PROCEEDINGS OF THE 2021 DESIGN, AUTOMATION & TEST IN EUROPE CONFERENCE & EXHIBITION (DATE 2021), 2021, : 1272 - 1277
  • [44] EODM: On Developing Enhanced Object Detection Model using Fast Region-based Convolution Neural Networks (FRCNN)
    Anuradha, B.
    Karthik, S.
    Mythili, S.
    Kavitha, M. S.
    TEHNICKI VJESNIK-TECHNICAL GAZETTE, 2024, 31 (02): : 566 - 573
  • [45] Asymmetric Convolution Networks Based on Multi-feature Fusion for Object Detection
    Yang, Zhenkun
    Ma, Xianghua
    An, Jing
    2020 IEEE 16TH INTERNATIONAL CONFERENCE ON AUTOMATION SCIENCE AND ENGINEERING (CASE), 2020, : 1355 - 1360
  • [46] SHetConv: target keypoint detection based on heterogeneous convolution neural networks
    Xiaojie Yin
    Ning He
    Xiaoxiao Liu
    Ke Lu
    Multimedia Systems, 2021, 27 : 519 - 529
  • [47] Vehicle Motion Detection Algorithm based on Novel Convolution Neural Networks
    Gao, Sheng-yang
    Jiang, Xian-yang
    Tang, Xiang-hong
    CURRENT TRENDS IN COMPUTER SCIENCE AND MECHANICAL AUTOMATION, VOL 1, 2017, : 544 - 556
  • [48] SHetConv: target keypoint detection based on heterogeneous convolution neural networks
    Yin, Xiaojie
    He, Ning
    Liu, Xiaoxiao
    Lu, Ke
    MULTIMEDIA SYSTEMS, 2021, 27 (03) : 519 - 529
  • [49] Cloud Detection and Tracking Based on Object Detection with Convolutional Neural Networks
    Carballo, Jose Antonio
    Bonilla, Javier
    Fernandez-Reche, Jesus
    Nouri, Bijan
    Avila-Marin, Antonio
    Fabel, Yann
    Alarcon-Padilla, Diego-Cesar
    ALGORITHMS, 2023, 16 (10)
  • [50] LOW-LATENCY LIGHTWEIGHT STREAMING SPEECH RECOGNITION WITH 8-BIT QUANTIZED SIMPLE GATED CONVOLUTIONAL NEURAL NETWORKS
    Park, Jinhwan
    Qian, Xue
    Jo, Youngmin
    Sung, Wonyong
    2020 IEEE INTERNATIONAL CONFERENCE ON ACOUSTICS, SPEECH, AND SIGNAL PROCESSING, 2020, : 1803 - 1807