AI Chat Paper
Note: Please note that the following content is generated by AMiner AI. SciOpen does not take any responsibility related to this content.
{{lang === 'zh_CN' ? '文章概述' : 'Summary'}}
{{lang === 'en_US' ? '中' : 'Eng'}}
Chat more with AI
PDF (1.4 MB)
Collect
Submit Manuscript AI Chat Paper
Show Outline
Outline
Show full outline
Hide outline
Outline
Show full outline
Hide outline
Publishing Language: Chinese | Open Access

Convolutional neural network mixed-precision quantization method considering layer sensitivity

Haijun LIUChenxi ZHANGXiyu WANGChanglin CHENJun CHENZhiwei LI( )
College of Electronic Science and Technology, National University of Defense Technology, Changsha 410073, China
Show Author Information

Abstract

To address the problem of how to faithfully map neural networks to resource-constrained embedded devices, a mixed-precision quantization method for convolutional neural networks based on layer sensitivity analysis was proposed. The sensitivity of convolutional layer parameters was measured by calculating the average trace of the Hessian matrix, providing a basis for bit-width allocation. A layer-wise ascending-descending approach was employed for bit-width allocation, ultimately achieving mixed-precision quantization of the network model. Experimental results demonstrate that compared to the fixed-precision quantization methods DoReFa and LSQ +, the proposed mixed-precision quantization method improves recognition accuracy by 10.2% and 1.7%, respectively, at an average bit-width of 3 bit. When compared to other mixed-precision quantization methods, the proposed approach achieves over 1% higher recognition accuracy. Additionally, noise-injected training effectively enhances the robustness of the mixed-precision quantization method, improving recognition accuracy by 16% under a noise standard deviation of 0.5.

CLC number: TP183 Document code: A Article ID: 1001-2486(2025)04-143-08

References

【1】
【1】
 
 
Journal of National University of Defense Technology
Pages 143-150

{{item.num}}

Comments on this article

Go to comment

< Back to all reports

Review Status: {{reviewData.commendedNum}} Commended , {{reviewData.revisionRequiredNum}} Revision Required , {{reviewData.notCommendedNum}} Not Commended Under Peer Review

Review Comment

Close
Close
Cite this article:
LIU H, ZHANG C, WANG X, et al. Convolutional neural network mixed-precision quantization method considering layer sensitivity. Journal of National University of Defense Technology, 2025, 47(4): 143-150. https://doi.org/10.11887/j.issn.1001-2486.25010015

1124

Views

16

Downloads

0

Crossref

0

Web of Science

0

Scopus

0

CSCD

Received: 10 January 2025
Published: 01 August 2025
© 2025 Journal of National University of Defense Technology

This is an open access article under the CC BY-NC-ND license (http://creativecommons.org/licenses/by-nc-nd/4.0/).