Efficient Compactions between Storage Tiers with PrismDB
Ashwini Raina, Jianan Lu, Asaf Cidon, Michael J. Freedman
Abstract
In recent years, emerging storage hardware technologies have focused on divergent goals: better performance or lower cost-per-bit. Correspondingly, data systems that employ these technologies are typically optimized either to be fast (but expensive) or cheap (but slow). We take a different approach: by architecting a storage engine to natively utilize two tiers of fast and low-cost storage technologies, we can achieve a Pareto-efficient balance between performance and cost-per-bit.
This paper presents the design and implementation of PrismDB, a novel key-value store that exploits two extreme ends of the spectrum of modern NVMe storage technologies (3D XPoint and QLC NAND) simultaneously. Our key contribution is how to efficiently migrate and compact data between two different storage tiers. Inspired by the classic cost-benefit analysis of log cleaning, we develop a new algorithm for multi-tiered storage compaction that balances the benefit of reclaiming space for hot objects in fast storage with the cost of compaction I/O in slow storage. Compared to the standard use of RocksDB on flash in datacenters today, PrismDB's average throughput on tiered storage is 3.3× faster and its read tail latency is 2× better, using equivalently-priced hardware.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0a55bc23-9e2a-4de4-bac5-da9a204a8affCited by top-tier papers6
- AegonKV: A High Bandwidth, Low Tail Latency, and Low Storage Cost KV-Separated LSM Store with SmartSSD-based GC OffloadingZhuohui Duan, Hao Feng, Haikun Liu, Xiaofei Liao et al.FAST 2025 · 11 citations
- HotRAP: Hot Record Retention and Promotion for LSM-trees with Tiered StorageJiansheng Qiu, Fangzhou Yuan, Mingyu Gao, Huanchen ZhangUSENIX ATC 2025 · 3 citations
- Mitigating Resource Usage Dependency in Sorting-based KV Stores on Hybrid Storage Devices via Operation DecouplingQingyang Zhang, Yongkun Li, Yubiao Pan, Haoting Tang et al.USENIX ATC 2025 · 3 citations
- Getting the MOST out of your Storage Hierarchy with Mirror-Optimized Storage TieringKaiwei Tu, Kan Wu, Andrea C. Arpaci-Dusseau, Remzi H. Arpaci-DusseauFAST 2026 · 2 citations
- PrvTel: Lightweight Models for Private and Accurate Telemetry Data RetentionYajie Zhou, Fuheng Zhao, Eric S. Wang, Ayse K. Coskun et al.NSDI 2026 · 1 citation
Builds on9
- A large scale analysis of hundreds of in-memory cache clusters at TwitterJuncheng Yang, Yao Yue, K. V. RashmiOSDI 2020 · 245 citations
- Learning Relaxed Belady for Content Distribution Network CachingZhenyu Song, Daniel S. Berger, Kai Li, Wyatt LloydNSDI 2020 · 193 citations
- MatrixKV: Reducing Write Stalls and Write Amplification in LSM-tree Based KV Stores with Matrix Container in NVMTing Yao, Yiwen Zhang, Jiguang Wan, Qiu Cui et al.USENIX ATC 2020 · 186 citations
- SpanDB: A Fast, Cost-Effective LSM-tree Based KV Store on Hybrid StorageHao Chen, Chaoyi Ruan, Cheng Li, Xiaosong Ma et al.FAST 2021 · 120 citations
- SplinterDB: Closing the Bandwidth Gap for NVMe Key-Value StoresAlexander Conway, Abhishek Gupta, Vijay Chidambaram, Martin Farach-Colton et al.USENIX ATC 2020 · 90 citations
Related papers
- Prism: Optimizing Key-Value Store for Modern Heterogeneous Storage DevicesYongju Song, Wook-Hee Kim, Sumit Kumar Monga, Changwoo Min et al.ASPLOS 2023 · 28 citations
- ListDB: Union of Write-Ahead Logs and Persistent SkipLists for Incremental Checkpointing on Persistent MemoryWonbae Kim, Chanyeol Park, Dongui Kim, Hyeongjun Park et al.OSDI 2022 · 47 citations
- PartitionKV: Redesigning LSM-tree KV Stores on NVMs with Adaptive Partitioning for Reducing Write Stalls and AmplificationXingye Huang, Jinyu Wu, Xiaofang Xia, Jiangtao Cui et al.SIGMOD 2026
- MirrorKV: An Efficient Key-Value Store on Hybrid Cloud Storage with Balanced Performance of Compaction and QueryingZhiqi Wang, Zili ShaoSIGMOD 2024 · 6 citations
- Spitfire: A Three-Tier Buffer Manager for Volatile and Non-Volatile MemoryXinjing Zhou, Joy Arulraj, Andrew Pavlo, David E. CohenSIGMOD 2021 · 42 citations
