Effective Mimicry of Belady's MIN Policy
Ishan Shah, Akanksha Jain, Calvin Lin
摘要
The past decade has seen the rise of highly successful cache replacement policies that are based on binary prediction. For example, the Hawkeye policy learns whether lines loaded by a given PC are Cache Friendly (likely to remain in the cache if Belady’s MIN policy had been used) or Cache Averse (likely to be evicted by Belady’s MIN policy). In this paper, we instead present a cache replacement policy that is based on multiclass prediction, which allows it to directly mimic Belady’s MIN policy in a surprisingly simple and effective way. Our policy uses a PC-based predictor to learn each cache line’s reuse distance; it then evicts lines based on their predicted time of reuse. We show that our use of multiclass prediction is more effective than binary prediction because it allows for a finer-grained ordering of cache lines during eviction and because it is more robust to prediction errors.Our empirical results show that our new policy, which we refer to as Mockingjay, outperforms the previous state-of-the-art on both single-core and multi-core platforms and both with and without a prefetcher. For example, with no prefetcher, on a mix of 100 multi-core workloads from the SPEC 2006, SPEC 2017, and GAP benchmark suites, Mockingjay sees an average improvement over LRU of 15.2%, compared to 7.6% for SHiP and 12.9% for Hawkeye. On a single-core platform, Mockingjay’s improvement over LRU is 5.7%, which approaches the 6.0% improvement of Belady MIN’s unrealizable policy. On a single-core platform (with a prefetcher) running the high-MPKI CVP workloads, Mockingjay’s improvement over LRU is 20.1%, compared to 13.4% for Hawkeye.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper22
- Triangel: A High-Performance, Accurate, Timely On-Chip Temporal PrefetcherSam Ainsworth, Lev MukhanovISCA 2024 · 被引用 22 次
- CLIP: Load Criticality based Data Prefetching for Bandwidth-constrained Many-core SystemsBiswabandan PandaMICRO 2023 · 被引用 21 次
- The Maya Cache: A Storage-efficient and Secure Fully-associative Last-level CacheAnubhav Bhatla, Navneet, Biswabandan PandaISCA 2024 · 被引用 12 次
- CARE: A Concurrency-Aware Enhanced Lightweight Cache Management FrameworkXiaoyang Lu, Rujia Wang, Xian-He SunHPCA 2023 · 被引用 11 次
- Constable: Improving Performance and Power Efficiency by Safely Eliminating Load Instruction ExecutionRahul Bera, Adithya Ranganathan, Joydeep Rakshit, Sujit Mahto 等ISCA 2024 · 被引用 8 次
它引用的顶会 Paper1
相关 Paper
- Light-weight Cache Replacement for Instruction Heavy WorkloadsSaba Mostofi, Setu Gupta, Ahmad Hassani, Krishnam Tibrewala 等ISCA 2025 · 被引用 6 次
- Drishti: Do Not Forget Slicing While Designing Last-Level Cache Replacement Policies for Many-Core SystemsSweta, Prerna Priyadarshini, Biswabandan PandaMICRO 2025 · 被引用 1 次
- SS-LRU: a smart segmented LRU cachingChunhua Li, Man Wu, Yuhan Liu, Ke Zhou 等DAC 2022 · 被引用 8 次
- R-Max: Extending BéLáDy's MIN with Prefetching to Bound Realistic Cache PerformanceLei Wang, Chia-Hang Lee, Maccoy Merrell, Gino Chacon 等ISCA 2026
- Reducing Load Latency with Cache Level PredictionMajid Jalili, Mattan ErezHPCA 2022 · 被引用 17 次
