AI Chat Paper
Note: Please note that the following content is generated by AMiner AI. SciOpen does not take any responsibility related to this content.
{{lang === 'zh_CN' ? '文章概述' : 'Summary'}}
{{lang === 'en_US' ? '中' : 'Eng'}}
Chat more with AI
PDF (3.3 MB)
Collect
Submit Manuscript AI Chat Paper
Show Outline
Outline
Show full outline
Hide outline
Outline
Show full outline
Hide outline
Publishing Language: Chinese | Open Access

High-efficiency data loading and output buffering strategy for sparse convolutional computing

Biao LIUChanglin CHENYufei ZHANGSitong LIULiqin TANGHongqi YU( )
College of Electronic Science and Technology, National University of Defense Technology, Changsha 410073, China
Show Author Information

Abstract

In view of the problems such as inefficient data loading, insufficient utilization of multiply-accumulates resources, complex output buffering and addressing logic in existing neural network accelerators when processing sparse neural networks, a high-efficiency data loading and output buffering strategy for sparse convolutional computing was proposed. It performed an all-to-all multiply-accumulates operation on the non-zero input feature map data and the non-zero weights belonging to the same input channel, which reduces the difficulty of non-zero data pairing and improves the utilization of multiply-accumulates resources. By using input stationary calculation and intensive cyclic loading of input feature map data, it significantly reduced the number of data off-chip fetches. It optimized the output buffer design and solved the problems of address access contention and storage congestion during output buffering in existing solutions. Experimental results show that, when compare to fine-grained systolic accelerator with similar architectures, the process element area of the proposed architecture is decreased by 21.45%; the data loading speed is increased by 117.71% on average; the average utilization of multiplier is increased by 11.25%, reaching 89%.

CLC number: TN492 Document code: A Article ID: 1001-2486(2023)05-212-10

References

【1】
【1】
 
 
Journal of National University of Defense Technology
Pages 212-221

{{item.num}}

Comments on this article

Go to comment

< Back to all reports

Review Status: {{reviewData.commendedNum}} Commended , {{reviewData.revisionRequiredNum}} Revision Required , {{reviewData.notCommendedNum}} Not Commended Under Peer Review

Review Comment

Close
Close
Cite this article:
LIU B, CHEN C, ZHANG Y, et al. High-efficiency data loading and output buffering strategy for sparse convolutional computing. Journal of National University of Defense Technology, 2023, 45(5): 212-221. https://doi.org/10.11887/j.cn.202305025

293

Views

1

Downloads

0

Crossref

0

Web of Science

0

Scopus

0

CSCD

Received: 08 June 2022
Published: 28 October 2023
© 2023 Journal of National University of Defense Technology

This is an open access article under the CC BY-NC-ND license (http://creativecommons.org/licenses/by-nc-nd/4.0/).