HALP: Heuristic Aided Learned Preference Eviction Policy for YouTube Content Delivery Network
Zhenyu Song, Kevin Chen, Nuikhil Sarda, Deniz Altinbüken, Eugene Brevdo, Jimmy Coleman, Xiao Ju, Pawel Jurczyk, Richard Schooler, Ramki Gummadi
摘要
Video streaming services are among the largest web applications in production, and a large source of downstream internet traffic. A large-scale video streaming service at Google, YouTube, leverages a Content Delivery Network (CDN) to serve its users. A key consideration in providing a seamless service is cache efficiency. In this work, we demonstrate machine learning techniques to improve the efficiency of YouTube's CDN DRAM cache. While many recently proposed learning-based caching algorithms show promising results, we identify and address three challenges blocking deployment of such techniques in a large-scale production environment: computation overhead for learning, robust byte miss ratio improvement, and measuring impact under production noise. We propose a novel caching algorithm, HALP, which achieves low CPU overhead and robust byte miss ratio improvement by augmenting a heuristic policy with machine learning. We also propose a production measurement method, impact distribution analysis, that can accurately measure the impact distribution of a new caching algorithm deployment in a noisy production environment.
HALP has been running in YouTube CDN production as a DRAM level eviction algorithm since early 2022 and has reliably reduced the byte miss during peak by an average of 9.1% while expending a modest CPU overhead of 1.8%.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- Baleen: ML Admission & Prefetching for Flash CachesDaniel Lin-Kit Wong, Hao Wu, Carson Molder, Sathya Gunasekar 等FAST 2024 · 被引用 26 次
- 3L-Cache: Low Overhead and Precise Learning-based Eviction Policy for CachesWenbin Zhou, Zhixiong Niu, Yongqiang Xiong, Juan Fang 等FAST 2025 · 被引用 16 次
- Seer: Enabling Future-Aware Online Caching in Networked SystemsJason Lei, Vishal ShrivastavNSDI 2024 · 被引用 11 次
- Learned Prefix Caching for Efficient LLM InferenceDongsheng Yang, Austin T. Li, Kai Li, Wyatt LloydNeurIPS 2025 · 被引用 8 次
- Robustifying Learning-Augmented Caching Efficiently without Compromising 1-ConsistencyPeng Chen, Hailiang Zhao, Jiaji Zhang, Xueyan Tang 等NeurIPS 2025 · 被引用 4 次
它引用的顶会 Paper4
- Learning Relaxed Belady for Content Distribution Network CachingZhenyu Song, Daniel S. Berger, Kai Li, Wyatt LloydNSDI 2020 · 被引用 193 次
- The CacheLib Caching Engine: Design and Experiences at ScaleBenjamin Berg, Daniel S. Berger, Sara McAllister, Isaac Grosof 等OSDI 2020 · 被引用 145 次
- An Imitation Learning Approach for Cache ReplacementEvan Zheran Liu, Milad Hashemi, Kevin Swersky, Parthasarathy Ranganathan 等ICML 2020 · 被引用 108 次
- Learning Cache Replacement with CACHEUSLiana V. Rodriguez, Farzana Beente Yusuf, Steven Lyons, Eysler Paz 等FAST 2021 · 被引用 26 次
相关 Paper
- RL-Bélády: A Unified Learning Framework for Content CachingGang Yan, Jian LiACM MM 2020 · 被引用 15 次
- Intelligent Video Caching at Network Edge: A Multi-Agent Deep Reinforcement Learning ApproachFangxin Wang, Feng Wang, Jiangchuan Liu, Ryan Shea 等INFOCOM 2020 · 被引用 139 次
- GL-Cache: Group-level learning for efficient and high-performance cachingJuncheng Yang, Ziming Mao, Yao Yue, K. V. RashmiFAST 2023 · 被引用 60 次
- Darwin: Flexible Learning-based CDN CachingJiayi Chen, Nihal Sharma, Tarannum Khan, Shu Liu 等SIGCOMM 2023 · 被引用 13 次
- Robust Learning-Augmented Caching: An Experimental StudyJakub Chledowski, Adam Polak, Bartosz Szabucki, Konrad Tomasz ZolnaICML 2021 · 被引用 21 次
