Catalyst: Optimizing Cache Management for Large In-memory Key-value Systems
Kefei Wang, Feng Chen
摘要
In-memory key-value cache systems, such as Memcached and Redis, are essential in today's data centers. A key mission of such cache systems is to identify the most valuable data for caching. To achieve this, the current system design keeps track of each key-value item's access and attempts to make accurate estimation on its temporal locality. All it aims is to achieve the highest cache hit ratio. However, as cache capacity quickly increases, the overhead of managing metadata for a massive amount of small key-value items rises to an unbearable level. Put it simply, the current fine-grained, heavy-cost approach cannot continue to scale.
In this paper, we have performed an experimental study on the scalability challenge of the current key-value cache system design and quantitatively analyzed the inherent issues related to the metadata operations for cache management. We further propose a key-value cache management scheme, called Catalyst , based on a highly efficient metadata structure, which allows us to make effective caching decisions in a scalable way. By offloading non-essential metadata operations to GPU, we can further dedicate the limited CPU and memory resources to the main service operations for improved throughput and latency. We have developed a prototype based on Memcached. Our experimental results show that our scheme can significantly enhance the scalability and improve the cache system performance by a factor of up to 4.3.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Ada-KV: Optimizing KV Cache Eviction by Adaptive Budget Allocation for Efficient LLM InferenceYuan Feng, Junlin Lv, Yukun Cao, Xike Xie 等NeurIPS 2025 · 被引用 256 次
- Cache-Craft: Managing Chunk-Caches for Efficient Retrieval-Augmented GenerationShubham Agarwal, Sai Sundaresan, Subrata Mitra, Debabrata Mahapatra 等SIGMOD 2025 · 被引用 20 次
- Mnemosyne: Dynamic Workload-Aware BF Tuning via Accurate Statistics in LSM treesZichen Zhu, Yanpeng Wei, Ju Hyoung Mun, Manos AthanassoulisSIGMOD 2025 · 被引用 3 次
- Efficient Cooperation-Aware Key and Value Management for LLM InferenceQiheng Sun, Hongwei Zhang, Junxu Liu, Haocheng Xia 等VLDB 2026
- Merlin: An Efficient Adaptive Cache Eviction Algorithm via Fine-Grained CharacterizationLiujia Li, Jinhao Guo, Yi Fan, Jianyu Wu 等OSDI 2026
它引用的顶会 Paper7
- FPGA-Accelerated Compactions for LSM-based Key-Value StoreTeng Zhang, Jianying Wang, Xuntao Cheng, Hao Xu 等FAST 2020 · 被引用 99 次
- Viper: An Efficient Hybrid PMem-DRAM Key-Value StoreLawrence Benson, Hendrik Makait, Tilmann RablVLDB 2021 · 被引用 86 次
- AC-Key: Adaptive Caching for LSM-based Key-Value StoresFenggang Wu, Ming-Hong Yang, Baoquan Zhang, David H. C. DuUSENIX ATC 2020 · 被引用 81 次
- HotRing: A Hotspot-Aware In-Memory Key-Value StoreJiqiang Chen, Liang Chen, Sheng Wang, Guoyun Zhu 等FAST 2020 · 被引用 80 次
- BMC: Accelerating Memcached using Safe In-kernel Caching and Pre-stack ProcessingYoann Ghigoff, Julien Sopena, Kahina Lazri, Antoine Blin 等NSDI 2021 · 被引用 79 次
相关 Paper
- Put an Elephant into a Fridge: Optimizing Cache Efficiency for In-memory Key-value StoresKefei Wang, Jian Liu, Feng ChenVLDB 2020 · 被引用 23 次
- Segcache: a memory-efficient and scalable in-memory key-value cache for small objectsJuncheng Yang, Yao Yue, Rashmi VinayakNSDI 2021 · 被引用 70 次
- ScalaCache: Scalable User-Space Page Cache Management with Software-Hardware CoordinationLi Peng, Yuda An, You Zhou, Chenxi Wang 等USENIX ATC 2024 · 被引用 6 次
- IceCache: Memory-Efficient KV-cache Management for Long-Sequence LLMsYuzhen Mao, Qitong Wang, Martin Ester, Ke LiICLR 2026 · 被引用 6 次
- FrozenHot Cache: Rethinking Cache Management for Modern HardwareZiyue Qiu, Juncheng Yang, Juncheng Zhang, Cheng Li 等EuroSys 2023 · 被引用 33 次
