Segcache: a memory-efficient and scalable in-memory key-value cache for small objects
Juncheng Yang, Yao Yue, Rashmi Vinayak
Abstract
Modern web applications heavily rely on in-memory keyvalue caches to deliver low-latency, high-throughput services. In-memory caches store small objects of size in the range of 10s to 1000s of bytes, and use TTLs widely for data freshness and implicit delete. Current solutions have relatively large per-object metadata and cannot remove expired objects promptly without incurring a high overhead. We present Segcache, which uses a segment-structured design that stores data in fixed-size segments with three key features: (1) it groups objects with similar creation and expiration time into the segments for efficient expiration and eviction, (2) it approximates some and lifts most per-object metadata into the shared segment header and shared information slot in the hash table for object metadata reduction, and (3) it performs segment-level bulk expiration and eviction with tiny critical sections for high scalability. Evaluation using production traces shows that Segcache uses 22-60% less memory than state-of-the-art designs for a variety of workloads. Segcache simultaneously delivers high throughput, up to 40% better than Memcached on a single thread. It exhibits close-to-linear scalability, providing a close to 8× speedup over Memcached with 24 threads.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4bef9d3a-dbf1-4e99-b2c2-516dce0d43f3Cited by top-tier papers23
- SIEVE is Simpler than LRU: an Efficient Turn-Key Eviction Algorithm for Web CachesYazhuo Zhang, Juncheng Yang, Yao Yue, Ymir Vigfusson et al.NSDI 2024 · 63 citations
- GL-Cache: Group-level learning for efficient and high-performance cachingJuncheng Yang, Ziming Mao, Yao Yue, K. V. RashmiFAST 2023 · 60 citations
- FIFO queues are all you need for cache evictionJuncheng Yang, Yazhuo Zhang, Ziyue Qiu, Yao Yue et al.SOSP 2023 · 54 citations
- Pacman: An Efficient Compaction Approach for Log-Structured Key-Value Store on Persistent MemoryJing Wang, Youyou Lu, Qing Wang, Minhui Xie et al.USENIX ATC 2022 · 44 citations
- Approximate Caching for Efficiently Serving Text-to-Image Diffusion ModelsShubham Agarwal, Subrata Mitra, Sarthak Chakraborty, Srikrishna Karanam et al.NSDI 2024 · 44 citations
Builds on3
- Learning Relaxed Belady for Content Distribution Network CachingZhenyu Song, Daniel S. Berger, Kai Li, Wyatt LloydNSDI 2020 · 193 citations
- FlatStore: An Efficient Log-Structured Key-Value Storage Engine for Persistent MemoryYoumin Chen, Youyou Lu, Fan Yang, Qing Wang et al.ASPLOS 2020 · 166 citations
- HotRing: A Hotspot-Aware In-Memory Key-Value StoreJiqiang Chen, Liang Chen, Sheng Wang, Guoyun Zhu et al.FAST 2020 · 80 citations
Related papers
- A large scale analysis of hundreds of in-memory cache clusters at TwitterJuncheng Yang, Yao Yue, K. V. RashmiOSDI 2020 · 245 citations
- Catalyst: Optimizing Cache Management for Large In-memory Key-value SystemsKefei Wang, Feng ChenVLDB 2023 · 6 citations
- AC-Cache: A Memory-Efficient Caching System for Small Objects via Exploiting Access CorrelationsFulin Nan, Ronglong Wu, Zhirong Shen, Jiahui Yang et al.PPoPP 2025 · 1 citation
- DLHT: A Non-blocking Resizable Hashtable with Fast Deletes and Memory-awarenessAntonios Katsarakis, Vasilis Gavrielatos, Nikos NtarmosHPDC 2024 · 3 citations
- STsCache: An Efficient Semantic Caching Scheme for Time-series Data Workloads Based on Hybrid StorageTao Kong, Hui Li, Yuxuan Zhao, Liping Li et al.VLDB 2025
