AI Chat Paper
Note: Please note that the following content is generated by AMiner AI. SciOpen does not take any responsibility related to this content.
{{lang === 'zh_CN' ? '文章概述' : 'Summary'}}
{{lang === 'en_US' ? '中' : 'Eng'}}
Chat more with AI
Article Link
Collect
Submit Manuscript
Show Outline
Outline
Show full outline
Hide outline
Outline
Show full outline
Hide outline
Regular Paper

P3DC: Reducing DRAM Cache Hit Latency by Hybrid Mappings

Ye Chi1,2,3,4,5Ren-Tong Guo1,2,3,4Xiao-Fei Liao1,2,3,4( )Hai-Kun Liu1,2,3,4Jianhui Yue6
National Engineering Research Center for Big Data Technology and System, Wuhan 430074, China
Services Computing Technology and System Laboratory, Wuhan 430074, China
Cluster and Grid Computing Laboratory, Wuhan 430074, China
School of Computer Science and Technology, Huazhong University of Science and Technology, Wuhan 430074, China
School of Big Data and Internet, Shenzhen Technology University, Shenzhen 518118, China
Department of Computer Science, Michigan Technological University, Houghton 49931-1295, U.S.A.
Show Author Information

Abstract

Die-stacked dynamic random access memory (DRAM) caches are increasingly advocated to bridge the performance gap between the on-chip cache and the main memory. To fully realize their potential, it is essential to improve DRAM cache hit rate and lower its cache hit latency. In order to take advantage of the high hit-rate of set-association and the low hit latency of direct-mapping at the same time, we propose a partial direct-mapped die-stacked DRAM cache called P3DC. This design is motivated by a key observation, i.e., applying a unified mapping policy to different types of blocks cannot achieve a high cache hit rate and low hit latency simultaneously. To address this problem, P3DC classifies data blocks into leading blocks and following blocks, and places them at static positions and dynamic positions, respectively, in a unified set-associative structure. We also propose a replacement policy to balance the miss penalty and the temporal locality of different blocks. In addition, P3DC provides a policy to mitigate cache thrashing due to block type variations. Experimental results demonstrate that P3DC can reduce the cache hit latency by 20.5% while achieving a similar cache hit rate compared with typical set-associative caches. P3DC improves the instructions per cycle (IPC) by up to 66% (12% on average) compared with the state-of-the-art direct-mapped cache—BEAR, and by up to 19% (6% on average) compared with the tag-data decoupled set-associative cache—DEC-A8.

Electronic Supplementary Material

Download File(s)
JCST-2206-12561-Highlights.pdf (546.4 KB)

References

【1】
【1】
 
 
Journal of Computer Science and Technology
Pages 1341-1360

{{item.num}}

Comments on this article

Go to comment

< Back to all reports

Review Status: {{reviewData.commendedNum}} Commended , {{reviewData.revisionRequiredNum}} Revision Required , {{reviewData.notCommendedNum}} Not Commended Under Peer Review

Review Comment

Close
Close
Cite this article:
Chi Y, Guo R-T, Liao X-F, et al. P3DC: Reducing DRAM Cache Hit Latency by Hybrid Mappings. Journal of Computer Science and Technology, 2024, 39(6): 1341-1360. https://doi.org/10.1007/s11390-023-2561-y

975

Views

0

Crossref

0

Web of Science

0

Scopus

0

CSCD

Received: 04 June 2022
Accepted: 28 September 2023
Published: 28 December 2024
© Institute of Computing Technology, Chinese Academy of Sciences 2024