Effective Mimicry of Belady's MIN Policy
Ishan Shah, Akanksha Jain, Calvin Lin
Abstract
The past decade has seen the rise of highly successful cache replacement policies that are based on binary prediction. For example, the Hawkeye policy learns whether lines loaded by a given PC are Cache Friendly (likely to remain in the cache if Belady’s MIN policy had been used) or Cache Averse (likely to be evicted by Belady’s MIN policy). In this paper, we instead present a cache replacement policy that is based on multiclass prediction, which allows it to directly mimic Belady’s MIN policy in a surprisingly simple and effective way. Our policy uses a PC-based predictor to learn each cache line’s reuse distance; it then evicts lines based on their predicted time of reuse. We show that our use of multiclass prediction is more effective than binary prediction because it allows for a finer-grained ordering of cache lines during eviction and because it is more robust to prediction errors.Our empirical results show that our new policy, which we refer to as Mockingjay, outperforms the previous state-of-the-art on both single-core and multi-core platforms and both with and without a prefetcher. For example, with no prefetcher, on a mix of 100 multi-core workloads from the SPEC 2006, SPEC 2017, and GAP benchmark suites, Mockingjay sees an average improvement over LRU of 15.2%, compared to 7.6% for SHiP and 12.9% for Hawkeye. On a single-core platform, Mockingjay’s improvement over LRU is 5.7%, which approaches the 6.0% improvement of Belady MIN’s unrealizable policy. On a single-core platform (with a prefetcher) running the high-MPKI CVP workloads, Mockingjay’s improvement over LRU is 20.1%, compared to 13.4% for Hawkeye.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f719ed7f-def0-409b-bf19-15971e5e0ecbCited by top-tier papers22
- Triangel: A High-Performance, Accurate, Timely On-Chip Temporal PrefetcherSam Ainsworth, Lev MukhanovISCA 2024 · 22 citations
- CLIP: Load Criticality based Data Prefetching for Bandwidth-constrained Many-core SystemsBiswabandan PandaMICRO 2023 · 21 citations
- The Maya Cache: A Storage-efficient and Secure Fully-associative Last-level CacheAnubhav Bhatla, Navneet, Biswabandan PandaISCA 2024 · 12 citations
- CARE: A Concurrency-Aware Enhanced Lightweight Cache Management FrameworkXiaoyang Lu, Rujia Wang, Xian-He SunHPCA 2023 · 11 citations
- Constable: Improving Performance and Power Efficiency by Safely Eliminating Load Instruction ExecutionRahul Bera, Adithya Ranganathan, Joydeep Rakshit, Sujit Mahto et al.ISCA 2024 · 8 citations
Builds on1
Related papers
- Light-weight Cache Replacement for Instruction Heavy WorkloadsSaba Mostofi, Setu Gupta, Ahmad Hassani, Krishnam Tibrewala et al.ISCA 2025 · 6 citations
- Drishti: Do Not Forget Slicing While Designing Last-Level Cache Replacement Policies for Many-Core SystemsSweta, Prerna Priyadarshini, Biswabandan PandaMICRO 2025 · 1 citation
- SS-LRU: a smart segmented LRU cachingChunhua Li, Man Wu, Yuhan Liu, Ke Zhou et al.DAC 2022 · 8 citations
- R-Max: Extending BéLáDy's MIN with Prefetching to Bound Realistic Cache PerformanceLei Wang, Chia-Hang Lee, Maccoy Merrell, Gino Chacon et al.ISCA 2026
- Reducing Load Latency with Cache Level PredictionMajid Jalili, Mattan ErezHPCA 2022 · 17 citations
