3L-Cache: Low Overhead and Precise Learning-based Eviction Policy for Caches
Wenbin Zhou, Zhixiong Niu, Yongqiang Xiong, Juan Fang, Qian Wang
摘要
Caches can effectively reduce request latency and network traffic, with the eviction policy serving as a core component. The effectiveness of an eviction policy is measured by both the byte miss ratio and the object miss ratio. To reduce these miss ratios, various learning-based policies have been proposed. However, the substantial computation overhead introduced by learning limits their deployment in production systems.
This work presents 3L-Cache, an object-level learning policy with Low computation overhead, while achieving the Lowest object miss ratio and the Lowest byte miss ratio among learning-based policies. To reduce overhead, we introduce two key advancements. First, we propose an efficient training data collection scheme that filters out unnecessary historical cache requests and dynamically adjusts the training frequency without compromising accuracy. Second, we design a low-overhead eviction method that integrates a bidirectional sampling policy to prioritize unpopular objects and an efficient eviction strategy to effectively select evicted objects. Furthermore, we incorporate a parameter auto-tuning method to enhance adaptability across traces.
We evaluate 3L-Cache in a testbed using 4855 traces. The results show that 3L-Cache reduces the average CPU overhead by 60.9% compared to HALP and by 94.9% compared to LRB. Additionally, 3L-Cache incurs only 6.4× the average overhead of LRU for small cache sizes and 3.4× for large cache sizes, while achieving the best byte miss ratio or object miss ratio among twelve state-of-the-art policies.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Learning-Augmented Heuristics: Simple Yet Smart, Robust and Interpretable Cache EvictionHaocheng Xia, William Nixon, Bintang Dwi Marthen, Pranav Bhandari 等OSDI 2026 · 被引用 1 次
- Nemo: A Low-Write-Amplification Cache for Tiny Objects on Log-Structured Flash DevicesXufeng Yang, Tingting Tan, Jingxin Hu, Congming Gao 等ASPLOS 2026
它引用的顶会 Paper11
- Learning Relaxed Belady for Content Distribution Network CachingZhenyu Song, Daniel S. Berger, Kai Li, Wyatt LloydNSDI 2020 · 被引用 193 次
- An Imitation Learning Approach for Cache ReplacementEvan Zheran Liu, Milad Hashemi, Kevin Swersky, Parthasarathy Ranganathan 等ICML 2020 · 被引用 108 次
- Segcache: a memory-efficient and scalable in-memory key-value cache for small objectsJuncheng Yang, Yao Yue, Rashmi VinayakNSDI 2021 · 被引用 70 次
- SIEVE is Simpler than LRU: an Efficient Turn-Key Eviction Algorithm for Web CachesYazhuo Zhang, Juncheng Yang, Yao Yue, Ymir Vigfusson 等NSDI 2024 · 被引用 63 次
- Fresh Caching for Dynamic ContentBahman Abolhassani, John Tadrous, Atilla Eryilmaz, Edmund YehINFOCOM 2021 · 被引用 61 次
相关 Paper
- GL-Cache: Group-level learning for efficient and high-performance cachingJuncheng Yang, Ziming Mao, Yao Yue, K. V. RashmiFAST 2023 · 被引用 60 次
- HALP: Heuristic Aided Learned Preference Eviction Policy for YouTube Content Delivery NetworkZhenyu Song, Kevin Chen, Nuikhil Sarda, Deniz Altinbüken 等NSDI 2023 · 被引用 36 次
- Designing a Cost-Effective Cache Replacement Policy using Machine LearningSubhash Sethumurugan, Jieming Yin, John SartoriHPCA 2021 · 被引用 81 次
- LBSC: A Cost-Aware Caching Framework for Cloud DatabasesZhaoxuan Ji, Zhongle Xie, Yuncheng Wu, Meihui ZhangICDE 2024 · 被引用 7 次
- Learned Prefix Caching for Efficient LLM InferenceDongsheng Yang, Austin T. Li, Kai Li, Wyatt LloydNeurIPS 2025 · 被引用 8 次
