FrozenHot Cache: Rethinking Cache Management for Modern Hardware
Ziyue Qiu, Juncheng Yang, Juncheng Zhang, Cheng Li, Xiaosong Ma, Qi Chen, Mao Yang, Yinlong Xu
Abstract
Caching is crucial for accelerating data access, employed as a ubiquitous design in modern systems at many parts of computer systems. With increasing core count, and shrinking latency gap between cache and modern storage devices, hit-path scalability becomes increasingly critical. However, existing production in-memory caches often use list-based management with promotion on each cache hit, which requires extensive locking and poses a significant overhead for scaling beyond a few cores. Moreover, existing techniques for improving scalability either (1) only focus on the indexing structure and do not improve cache management scalability, or (2) sacrifice efficiency or miss-path scalability.
Inspired by highly skewed data popularity and short-term hotspot stability in cache workloads, we propose Frozen-Hot, a generic approach to improve the scalability of listbased caches. FrozenHot partitions the cache space into two parts: a frozen cache and a dynamic cache. The frozen cache serves requests for hot objects with minimal latency by eliminating promotion and locking, while the latter leverages the existing cache design to achieve workload adaptivity. We built FrozenHot as a library that can be easily integrated into existing systems. We demonstrate its performance by enabling FrozenHot in two production systems: HHVM and RocksDB using under 100 lines of code. Evaluated using production traces from MSR and Twitter, FrozenHot improves the throughput of three baseline cache algorithms by up to 551%. Compared to stock RocksDB, FrozenHot-enhanced RocksDB shows a higher throughput on all YCSB workloads with up to 90% increase, as well as reduced tail latency.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8eaedeaa-b2de-4eff-88a3-dfc4ce8d653dCited by top-tier papers9
- FIFO queues are all you need for cache evictionJuncheng Yang, Yazhuo Zhang, Ziyue Qiu, Yao Yue et al.SOSP 2023 · 54 citations
- PIMLex: A High-Performance Learned Index with Processing-in-MemoryLixiao Cui, Kedi Yang, Yusen Li, Gang Wang et al.FAST 2025 · 11 citations
- Reducing Cross-Cloud/Region Costs with the Auto-Configuring MACARON CacheHojin Park, Ziyue Qiu, Gregory R. Ganger, George AmvrosiadisSOSP 2024 · 3 citations
- GPHash: An Efficient Hash Index for GPU with Byte-Granularity Persistent MemoryMenglei Chen, Yu Hua, Zhangyu Chen, Ming Zhang et al.FAST 2025 · 3 citations
- Getting the MOST out of your Storage Hierarchy with Mirror-Optimized Storage TieringKaiwei Tu, Kan Wu, Andrea C. Arpaci-Dusseau, Remzi H. Arpaci-DusseauFAST 2026 · 2 citations
Builds on11
- A large scale analysis of hundreds of in-memory cache clusters at TwitterJuncheng Yang, Yao Yue, K. V. RashmiOSDI 2020 · 245 citations
- Learning Relaxed Belady for Content Distribution Network CachingZhenyu Song, Daniel S. Berger, Kai Li, Wyatt LloydNSDI 2020 · 193 citations
- The CacheLib Caching Engine: Design and Experiences at ScaleBenjamin Berg, Daniel S. Berger, Sara McAllister, Isaac Grosof et al.OSDI 2020 · 145 citations
- SpanDB: A Fast, Cost-Effective LSM-tree Based KV Store on Hybrid StorageHao Chen, Chaoyi Ruan, Cheng Li, Xiaosong Ma et al.FAST 2021 · 120 citations
- HotRing: A Hotspot-Aware In-Memory Key-Value StoreJiqiang Chen, Liang Chen, Sheng Wang, Guoyun Zhu et al.FAST 2020 · 80 citations
Related papers
- twCache: Thread-Wise Cache Management with High Concurrency PerformanceYigui Yuan, Peiquan Jin, Xiaoliang WangICDE 2025 · 2 citations
- Fairer and More Scalable Reader-Writer Locks by Optimizing Queue ManagementTakashi Hoshino, Kenjiro TauraPPoPP 2025 · 1 citation
- HotPrefix: Hotness-Aware KV Cache Scheduling for Efficient Prefix Sharing in LLM Inference SystemsYuhang Li, Rong Gu, Chengying Huan, Zhibin Wang et al.SIGMOD 2026 · 9 citations
- AC-Cache: A Memory-Efficient Caching System for Small Objects via Exploiting Access CorrelationsFulin Nan, Ronglong Wu, Zhirong Shen, Jiahui Yang et al.PPoPP 2025 · 1 citation
- HotRAP: Hot Record Retention and Promotion for LSM-trees with Tiered StorageJiansheng Qiu, Fangzhou Yuan, Mingyu Gao, Huanchen ZhangUSENIX ATC 2025 · 3 citations
