Efficient Caching with A Tag-enhanced DRAM
Maryam Babaie, Ayaz Akram, Wendy Elsasser, Brent Haukness, Michael R. Miller, Taeksang Song, Thomas Vogelsang, Steven C. Woo, Jason Lowe-Power
摘要
As SRAM-based caches are hitting a scaling wall, manufacturers are integrating DRAM-based caches into system designs to continue increasing cache sizes. While DRAM caches can improve the performance of memory systems, existing DRAM cache designs suffer from high miss penalties, wasted data movement, and interference between misses and demands. In this paper, we propose TDRAM, a novel DRAM microarchitecture tailored for caching. TDRAM enhances existing DRAM, such as HBM3, by adding small, low-latency mats to store tags and metadata on the same die as the data mats. These mats enable tag and data access in lockstep, in-DRAM tag comparison, and conditional data response based on the comparison result (reducing wasted data transfers), akin to SRAM cache mechanisms. TDRAM further optimizes hit and miss latencies through opportunistic early tag probing. Moreover, TDRAM introduces a flush buffer to store conflicting dirty data on write misses, eliminating data bus turnaround delays on write demands. We evaluate TDRAM in a full-system simulation using a set of HPC workloads with large memory footprints, showing that TDRAM, on average, provides faster tag checks, speedup, and 21% less energy consumption compared to state-of-the-art commercial and research designs.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper6
- HeMem: Scalable Tiered Memory Management for Big Data Applications and Real NVMAmanda Raybuck, Tim Stamler, Wei Zhang, Mattan Erez 等SOSP 2021 · 被引用 93 次
- FIGARO: Improving System Performance via Fine-Grained In-DRAM Data Relocation and CachingYaohua Wang, Lois Orosa, Xiangjun Peng, Yang Guo 等MICRO 2020 · 被引用 72 次
- Bandwidth-Effective DRAM Cache for GPU s with Storage-Class MemoryJeongmin Hong, Sungjun Cho, Geonwoo Park, Wonhyuk Yang 等HPCA 2024 · 被引用 21 次
- LoopPoint: Checkpoint-driven Sampled Simulation for Multi-threaded ApplicationsAlen Sabu, Harish Patil, Wim Heirman, Trevor E. CarlsonHPCA 2022 · 被引用 21 次
- NOMAD: Enabling Non-blocking OS-managed DRAM Cache via Tag-Data DecouplingYoungin Kim, Hyeonjin Kim, William J. SongHPCA 2023 · 被引用 10 次
相关 Paper
- Native DRAM Cache: Re-architecting DRAM as a Large-Scale Cache for Data CentersYesin Ryu, Yoojin Kim, Giyong Jung, Jung Ho Ahn 等ISCA 2024 · 被引用 5 次
- Hybrid2: Combining Caching and Migration in Hybrid Memory SystemsEvangelos Vasilakis, Vassilis Papaefstathiou, Pedro Trancoso, Ioannis SourdisHPCA 2020 · 被引用 37 次
- Genie Cache: Non-Blocking Miss Handling and Replacement in Page-Table-Based DRAM CacheYoungin Kim, William J. SongMICRO 2024 · 被引用 3 次
- HUNTER: Releasing Persistent Memory Write Performance with A Novel PM-DRAM Collaboration ArchitectureYanqi Pan, Yifeng Zhang, Wen Xia, Xiangyu Zou 等DAC 2023 · 被引用 1 次
- BARD: Reducing Write Latency of DDR5 Memory by Exploiting Bank-ParallelismSuhas K. Vittal, Moinuddin QureshiHPCA 2026 · 被引用 2 次
