Leaper: A Learned Prefetcher for Cache Invalidation in LSM-tree based Storage Engines
Lei Yang, Hong Wu, Tieying Zhang, Xuntao Cheng, Feifei Li, Lei Zou, Yujie Wang, Rongyao Chen, Jianying Wang, Gui Huang
摘要
Frequency-based cache replacement policies that work well on page-based database storage engines are no longer sufficient for the emerging LSM-tree (Log-Structure Merge-tree) based storage engines. Due to the append-only and copyon-write techniques applied to accelerate writes, the stateof-the-art LSM-tree adopts mutable record blocks and issues frequent background operations (i.e., compaction, flush) to reorganize records in possibly every block. As a side-effect, such operations invalidate the corresponding entries in the cache for each involved record, causing sudden drops on the cache hit rates and spikes on access latency. Given the observation that existing methods cannot address this cache invalidation problem, we propose Leaper, a machine learning method to predict hot records in an LSM-tree storage engine and prefetch them into the cache without being disturbed by background operations. We implement Leaper in a state-of-the-art LSM-tree storage engine, X-Engine, as a light-weight plug-in. Evaluation results show that Leaper eliminates about 70% cache invalidations and 99% latency spikes with at most 0.95% overheads as measured in realworld workloads.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper18
- Updatable Learned Index with Precise PositionsJiacheng Wu, Yong Zhang, Shimin Chen, Yu Chen 等VLDB 2021 · 被引用 160 次
- Constructing and Analyzing the LSM Compaction Design SpaceSubhadeep Sarkar, Dimitris Staratzis, Zichen Zhu, Manos AthanassoulisVLDB 2021 · 被引用 73 次
- GL-Cache: Group-level learning for efficient and high-performance cachingJuncheng Yang, Ziming Mao, Yao Yue, K. V. RashmiFAST 2023 · 被引用 60 次
- Spooky: Granulating LSM-Tree Compactions CorrectlyNiv Dayan, Tamar Weiss, Shmuel Dashevsky, Michael Pan 等VLDB 2022 · 被引用 57 次
- Quantized Training of Gradient Boosting Decision TreesYu Shi, Guolin Ke, Zhuoming Chen, Shuxin Zheng 等NeurIPS 2022 · 被引用 51 次
它引用的顶会 Paper1
相关 Paper
- SA-LSM : Optimize Data Layout for LSM-tree Based Storage using Survival AnalysisTeng Zhang, Jian Tan, Xin Cai, Jianying Wang 等VLDB 2022 · 被引用 10 次
- FPGA-Accelerated Compactions for LSM-based Key-Value StoreTeng Zhang, Jianying Wang, Xuntao Cheng, Hao Xu 等FAST 2020 · 被引用 99 次
- LeaderKV: Improving Read Performance of KV Stores via Learned Index and Decoupled KV TableYi Wang, Jianan Yuan, Shangyu Wu, Huan Liu 等ICDE 2024 · 被引用 12 次
- Range Cache: An Efficient Cache Component for Accelerating Range Queries on LSM - Based Key-Value StoresXiaoliang Wang, Peiquan Jin, Yongping Luo, Zhaole ChuICDE 2024 · 被引用 10 次
- FPGA-based Compaction Engine for Accelerating LSM-tree Key-Value StoresXuan Sun, Jinghuan Yu, Zimeng Zhou, Chun Jason XueICDE 2020 · 被引用 34 次
