USENIX ATC2023顶会
Adaptive Online Cache Capacity Optimization via Lightweight Working Set Size Estimation at Scale
Rong Gu, Simian Li, Haipeng Dai, Hancheng Wang, Yili Luo, Bin Fan, Ran Ben Basat, Ke Wang, Zhenyu Song, Shouwei Chen, Beinan Wang, Yihua Huang, Guihai Chen
摘要
Big data applications extensively use cache techniques to accelerate data access. A key challenge for improving cache utilization is provisioning a suitable cache size to fit the dynamic working set size (WSS) and understanding the related item repetition ratio (IRR) of the trace. We propose Cuki, an approximate data structure for efficiently estimating online WSS and IRR for variable-size item access with proven accuracy guarantee. Our solution is cache-friendly, thread-safe, and light-weighted in design. Based on that, we design an adaptive online cache capacity tuning mechanism. Moreover, Cuki can also be adapted to accurately estimate the cache miss ratio curve (MRC) online. We built Cuki as a lightweight plugin of the widely-used distributed file caching system Alluxio. Evaluation results show that Cuki has higher accuracy than four state-of-the-art algorithms by over an order of magnitude and with better stability in performance. The end-to-end data access experiments show that the adaptive cache tuning framework using Cuki reduces the table querying latency by 79% and improves the file reading throughput by 29% on average. Compared with the cutting-edge MRC approach, Cuki uses less memory and improves accuracy by around 73% on average. Cuki is deployed on one of the world's largest social platforms to run the Presto query workloads.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Reducing Cross-Cloud/Region Costs with the Auto-Configuring MACARON CacheHojin Park, Ziyue Qiu, Gregory R. Ganger, George AmvrosiadisSOSP 2024 · 被引用 3 次
- FLOWS: Balanced MRC Profiling for Heterogeneous Object-Size CacheXiaojun Guo, Hua Wang, Ke Zhou, Hong Jiang 等EuroSys 2024 · 被引用 1 次
- JitterSketch: Finding Jittery Flows in Network StreamsZhongxian Liang, Qilong Shi, Xiyan Liang, Zihan Li 等WWW 2026
它引用的顶会 Paper13
- Autopilot: workload autoscaling at GoogleKrzysztof Rzadca, Pawel Findeisen, Jacek Swiderski, Przemyslaw Zych 等EuroSys 2020 · 被引用 299 次
- A large scale analysis of hundreds of in-memory cache clusters at TwitterJuncheng Yang, Yao Yue, K. V. RashmiOSDI 2020 · 被引用 245 次
- FaasCache: keeping serverless computing alive with greedy-dual cachingAlexander Fuerst, Prateek SharmaASPLOS 2021 · 被引用 223 次
- InfiniCache: Exploiting Ephemeral Serverless Functions to Build a Cost-Effective Memory CacheAo Wang, Jingyuan Zhang, Xiaolong Ma, Ali Anwar 等FAST 2020 · 被引用 118 次
- OFC: an opportunistic caching system for FaaS platformsDjob Mvondo, Mathieu Bacou, Kevin Nguetchouang, Lucien Ngale 等EuroSys 2021 · 被引用 80 次
相关 Paper
- RepBun: Load-Balanced, Shuffle-Free Cluster Caching for Structured DataMinchen Yu, Yinghao Yu, Yunchuan Zheng, Baichen Yang 等INFOCOM 2020 · 被引用 1 次
- TTLs Matter: Efficient Cache Sizing with TTL-Aware Miss Ratio Curves and Working Set SizesSari Sultan, Kia Shakiba, Albert Lee, Paul Chen 等EuroSys 2024 · 被引用 8 次
- OSCA: An Online-Model Based Cache Allocation Scheme in Cloud Block Storage SystemsYu Zhang, Ping Huang, Ke Zhou, Hua Wang 等USENIX ATC 2020 · 被引用 74 次
- Quantile Estimation with DuplicatesTianrui Xia, Ziling Chen, Shaoxu SongSIGMOD 2026
- LPCA: learned MRC profiling based cache allocation for file storage systemsYibin Gu, Yifan Li, Hua Wang, Li Liu 等DAC 2022 · 被引用 4 次
