AI Chat Paper
Note: Please note that the following content is generated by AMiner AI. SciOpen does not take any responsibility related to this content.
{{lang === 'zh_CN' ? '文章概述' : 'Summary'}}
{{lang === 'en_US' ? '中' : 'Eng'}}
Chat more with AI
PDF (741.1 KB)
Collect
Submit Manuscript AI Chat Paper
Show Outline
Outline
Show full outline
Hide outline
Outline
Show full outline
Hide outline
Publishing Language: Chinese | Open Access

The Hybrid Precision Convolutional Neural Network Accelerator Based on Block Floating Point

Xueang Yu, Jingsong Luo, Bosheng Liu, Jigang Wu( )
School of Computer Science and Technology, Guangdong University of Technology, Guangzhou 510006, Guangdong, China
Show Author Information

Abstract

With the continuous development of convolutional neural networks (CNNs) in deep learning, computational complexity and hardware resource consumption have become the significant bottlenecks limiting computational efficiency. This paper proposes a hybrid processing unit (HPE) based on Block Floating Point (BFP) , which optimizes the design of the convolution computation unit in the hardware architecture by replacing the traditional Look-up Table (LUT) with DSP and employing data packing techniques. This design enables flexible switching between INT4 and BFP8 computation modes, significantly improving computational performance and reducing hardware resource consumption. Experimental results show that, when using a hybrid precision (INT4 and BFP8) computation mode, HPE significantly reduces LUT and FF overhead, with hardware resource utilization efficiency increasing by 123.40% and 58.16%, respectively, compared to the baseline. Furthermore, the data packing techniques enable the HPE to achieve 2× higher throughput than the conventional implementations. This study provides an efficient hardware solution for deep learning acceleration, with broad potential applications, especially in deep learning tasks requiring high computational efficiency and resource optimization.

CLC number: TP391.4 Document code: A Article ID: 1007–7162(2026)5–118–7

References

【1】
【1】
 
 
Journal of Guangdong University of Technology
Pages 118-124

{{item.num}}

Comments on this article

Go to comment

< Back to all reports

Review Status: {{reviewData.commendedNum}} Commended , {{reviewData.revisionRequiredNum}} Revision Required , {{reviewData.notCommendedNum}} Not Commended Under Peer Review

Review Comment

Close
Close
Cite this article:
Yu X, Luo J, Liu B, et al. The Hybrid Precision Convolutional Neural Network Accelerator Based on Block Floating Point. Journal of Guangdong University of Technology, 2026, 43(5): 118-124. https://doi.org/10.12052/gdutxb.250056

1

Views

0

Downloads

0

Crossref

Received: 05 March 2025
Accepted: 22 April 2025
Published: 22 May 2025
© 2026 Editorial Office of Journal of Guangdong University of Technology

This is an open access article under the CC BY-NC-ND license (https://creativecommons.org/licenses/by-nc-nd/4.0/).