Learning Relaxed Belady for Content Distribution Network Caching
Zhenyu Song, Daniel S. Berger, Kai Li, Wyatt Lloyd
Abstract
This paper presents a new approach for caching in CDNs that uses machine learning to approximate the Belady MIN (oracle) algorithm. To accomplish this complex task, we propose a CDN cache design called Learning Relaxed Belady (LRB) to mimic a Relaxed Belady algorithm, using the concept of Belady boundary. We also propose a metric called good decision ratio to help us make better design decisions. In addition, the paper addresses several challenges to build an end-to-end machine learning caching prototype, including how to gather training data, limit memory overhead, and have lightweight training and prediction.
We have implemented an LRB simulator and a prototype within Apache Traffic Server. Our simulation results with 6 production CDN traces show that LRB reduces WAN traffic compared to a typical production CDN cache design by 4-25%, and consistently outperform other state-of-the-art methods. Our evaluation of the LRB prototype shows its overhead is modest and it can be deployed on today's CDN servers.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 22238b3c-f1cc-4452-b752-6472938923baCited by top-tier papers39
- The CacheLib Caching Engine: Design and Experiences at ScaleBenjamin Berg, Daniel S. Berger, Sara McAllister, Isaac Grosof et al.OSDI 2020 · 145 citations
- Segcache: a memory-efficient and scalable in-memory key-value cache for small objectsJuncheng Yang, Yao Yue, Rashmi VinayakNSDI 2021 · 70 citations
- GL-Cache: Group-level learning for efficient and high-performance cachingJuncheng Yang, Ziming Mao, Yao Yue, K. V. RashmiFAST 2023 · 60 citations
- LiveNet: a low-latency video transport network for large-scale live streamingJinyang Li, Zhenyu Li, Ri Lu, Kai Xiao et al.SIGCOMM 2022 · 59 citations
- Polyjuice: High-Performance Transactions via Learned Concurrency ControlJia-Chen Wang, Ding Ding, Huan Wang, Conrad Christensen et al.OSDI 2021 · 39 citations
Related papers
- RL-Bélády: A Unified Learning Framework for Content CachingGang Yan, Jian LiACM MM 2020 · 15 citations
- Caching with Delayed HitsNirav Atre, Justine Sherry, Weina Wang, Daniel S. BergerSIGCOMM 2020 · 47 citations
- Towards Latency Awareness for Content Delivery Network CachingGang Yan, Jian LiUSENIX ATC 2022 · 25 citations
- HALP: Heuristic Aided Learned Preference Eviction Policy for YouTube Content Delivery NetworkZhenyu Song, Kevin Chen, Nuikhil Sarda, Deniz Altinbüken et al.NSDI 2023 · 36 citations
- LBSC: A Cost-Aware Caching Framework for Cloud DatabasesZhaoxuan Ji, Zhongle Xie, Yuncheng Wu, Meihui ZhangICDE 2024 · 7 citations
