GL-Cache: Group-level learning for efficient and high-performance caching
Juncheng Yang, Ziming Mao, Yao Yue, K. V. Rashmi
Abstract
Web applications rely heavily on software caches to achieve low-latency, high-throughput services. To adapt to changing workloads, three types of learned caches (learned evictions) have been designed in recent years: object-level learning, learning-from-distribution, and learning-from-simple-experts. However, we argue that the learning granularity in existing approaches is either too fine (object-level), incurring significant computation and storage overheads, or too coarse (workload or expert-level) to capture the differences between objects and leaves a considerable efficiency gap.
In this work, we propose a new approach for learning in caches ("group-level learning"), which clusters similar objects into groups and performs learning and eviction at the group level. Learning at the group level accumulates more signals for learning, leverages more features with adaptive weights, and amortizes overheads over objects, thereby achieving both high efficiency and high throughput.
We designed and implemented GL-Cache on an opensource production cache to demonstrate group-level learning. Evaluations on 118 production block I/O and CDN cache traces show that GL-Cache has a higher hit ratio and higher throughput than state-of-the-art designs. Compared to LRB (object-level learning), GL-Cache improves throughput by 228× and hit ratio by 7% on average across cache sizes. For 10% of the traces (P90), GL-Cache provides a 25% hit ratio increase from LRB. Compared to the best of all learned caches, GL-Cache achieves a 64% higher throughput, a 3% higher hit ratio on average, and a 13% hit ratio increase at the P90.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 430857c5-6290-443a-b203-80073ed6f9a5Cited by top-tier papers17
- FIFO queues are all you need for cache evictionJuncheng Yang, Yazhuo Zhang, Ziyue Qiu, Yao Yue et al.SOSP 2023 · 54 citations
- FrozenHot Cache: Rethinking Cache Management for Modern HardwareZiyue Qiu, Juncheng Yang, Juncheng Zhang, Cheng Li et al.EuroSys 2023 · 33 citations
- 3L-Cache: Low Overhead and Precise Learning-based Eviction Policy for CachesWenbin Zhou, Zhixiong Niu, Yongqiang Xiong, Juan Fang et al.FAST 2025 · 16 citations
- Optimizing File Systems on Heterogeneous Memory by Integrating DRAM Cache with Virtual Memory ManagementYubo Liu, Yuxin Ren, Mingrui Liu, Hongbo Li et al.FAST 2024 · 13 citations
- Seer: Enabling Future-Aware Online Caching in Networked SystemsJason Lei, Vishal ShrivastavNSDI 2024 · 11 citations
Builds on21
- Lessons Learned from the Chameleon TestbedKate Keahey, Jason Anderson, Zhuo Zhen, Pierre Riteau et al.USENIX ATC 2020 · 398 citations
- ALEX: An Updatable Adaptive Learned IndexJialin Ding, Umar Farooq Minhas, Jia Yu, Chi Wang et al.SIGMOD 2020 · 274 citations
- A large scale analysis of hundreds of in-memory cache clusters at TwitterJuncheng Yang, Yao Yue, K. V. RashmiOSDI 2020 · 245 citations
- Bao: Making Learned Query Optimization PracticalRyan Marcus, Parimarjan Negi, Hongzi Mao, Nesime Tatbul et al.SIGMOD 2021 · 242 citations
- Learning Relaxed Belady for Content Distribution Network CachingZhenyu Song, Daniel S. Berger, Kai Li, Wyatt LloydNSDI 2020 · 193 citations
Related papers
- Segcache: a memory-efficient and scalable in-memory key-value cache for small objectsJuncheng Yang, Yao Yue, Rashmi VinayakNSDI 2021 · 70 citations
- Improving Range Scan Performance in LSM-trees with Group CachingHengrui Wang, Jiaoyi Zhang, Jiansheng Qiu, Fangzhou Yuan et al.SIGMOD 2026
- Merlin: An Efficient Adaptive Cache Eviction Algorithm via Fine-Grained CharacterizationLiujia Li, Jinhao Guo, Yi Fan, Jianyu Wu et al.OSDI 2026
- Darwin: Flexible Learning-based CDN CachingJiayi Chen, Nihal Sharma, Tarannum Khan, Shu Liu et al.SIGCOMM 2023 · 13 citations
- SIEVE is Simpler than LRU: an Efficient Turn-Key Eviction Algorithm for Web CachesYazhuo Zhang, Juncheng Yang, Yao Yue, Ymir Vigfusson et al.NSDI 2024 · 63 citations
