AI Chat Paper
Note: Please note that the following content is generated by AMiner AI. SciOpen does not take any responsibility related to this content.
{{lang === 'zh_CN' ? '文章概述' : 'Summary'}}
{{lang === 'en_US' ? '中' : 'Eng'}}
Chat more with AI
Article Link
Collect
Submit Manuscript
Show Outline
Outline
Show full outline
Hide outline
Outline
Show full outline
Hide outline
Regular Paper

LayCO: Achieving Least Lossy Accuracy for Most EfficientRRAM-Based Deep Neural Network Accelerator via Layer-Centric Co-Optimization

Shao-Feng Zhao1,2,3Fang Wang1,2,4( )Bo Liu5( )Dan Feng1,2Yang Liu3
Wuhan National Laboratory for Optoelectronics, Huazhong University of Science and Technology, Wuhan 430074, China
School of Computer Science and Technology, Huazhong University of Science and Technology, Wuhan 430074, China
Cloud Computing and Big Data Institute, Henan University of Economics and Law, Zhengzhou 450001, China
Research Institute of Huazhong University of Science and Technology in Shenzhen, Shenzhen 518057, China
School of Computer and Artificial Intelligence, Zhengzhou University, Zhengzhou 450001, China
Show Author Information

Abstract

Resistive random access memory (RRAM) enables the functionality of operating massively parallel dot products and accumulations. RRAM-based accelerator is such an effective approach to bridging the gap between Internet of Things devices’ constrained resources and deep neural networks’ tremendous cost. Due to the huge overhead of Analog to Digital (A/D) and digital accumulations, analog RRAM buffer is introduced to extend the processing in analog and in approximation. Although analog RRAM buffer offers potential solutions to A/D conversion issues, the energy consumption is still challenging in resource-constrained environments, especially with enormous intermediate data volume. Besides, critical concerns over endurance must also be resolved before the RRAM buffer could be frequently used in reality for DNN inference tasks. Then we propose LayCO, a layer-centric co-optimizing scheme to address the energy and endurance concerns altogether while strictly providing an inference accuracy guarantee. LayCO relies on two key ideas: 1) co-optimizing with reduced supply voltage and reduced bit-width of accelerator architectures to increase the DNN’s error tolerance and achieve the accelerator’s energy efficiency, and 2) efficiently mapping and swapping individual DNN data to a corresponding RRAM partition in a way that meets the endurance requirements. The evaluation with representative DNN models demonstrates that LayCO outperforms the baseline RRAM buffer based accelerator by 27x improvement in energy efficiency (over TIMELY-like configuration), 308x in lifetime prolongation and 6x in area reduction (over RAQ) while maintaining the DNN accuracy loss less than 1%.

Electronic Supplementary Material

Download File(s)
JCST-2205-12545-Highlights.pdf (102 KB)

References

【1】
【1】
 
 
Journal of Computer Science and Technology
Pages 328-347

{{item.num}}

Comments on this article

Go to comment

< Back to all reports

Review Status: {{reviewData.commendedNum}} Commended , {{reviewData.revisionRequiredNum}} Revision Required , {{reviewData.notCommendedNum}} Not Commended Under Peer Review

Review Comment

Close
Close
Cite this article:
Zhao S-F, Wang F, Liu B, et al. LayCO: Achieving Least Lossy Accuracy for Most EfficientRRAM-Based Deep Neural Network Accelerator via Layer-Centric Co-Optimization. Journal of Computer Science and Technology, 2023, 38(2): 328-347. https://doi.org/10.1007/s11390-023-2545-y

903

Views

1

Crossref

1

Web of Science

1

Scopus

0

CSCD

Received: 31 May 2022
Accepted: 26 March 2023
Published: 30 March 2023
© Institute of Computing Technology, Chinese Academy of Sciences 2023