Understanding the Performance Implications of the Design Principles in Storage-Disaggregated Databases
Xi Pang, Jianguo Wang
摘要
Storage-compute disaggregation has recently emerged as a novel architecture in modern data centers, particularly in the cloud. By decoupling compute from storage, this new architecture enables independent and elastic scaling of compute and storage resources, potentially increasing resource utilization and reducing overall costs. To best leverage the disaggregated architecture, a new breed of database systems termed storage-disaggregated databases has recently been developed, such as Amazon Aurora, Microsoft Socrates, Google AlloyDB, Alibaba PolarDB, and Huawei Taurus. However, little is known about the effectiveness of the design principles in these databases since they are typically developed by industry giants, and only the overall performance results are presented without detailing the impact of individual design principles. As a result, many critical research questions remain unclear, such as the performance impact of storage-disaggregation, the log-as-the-database design, shared-storage, and various log-replay methods.
In this paper, we investigate the performance implications of the design principles that are widely adopted in storage-disaggregated databases for the first time. As these databases were usually not open-sourced, we have made a significant effort to implement a storage-disaggregated database prototype based on PostgreSQL v13.0. By fully controlling and instrumenting the codebase, we are able to selectively enable and disable individual optimizations and techniques to evaluate their impact on performance in various scenarios. Furthermore, we open-source our storage-disaggregated database prototype for use by the broader database research community, fostering collaboration and innovation in this field.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- CaaS-LSM: Compaction-as-a-Service for LSM-based Key-Value Stores in Storage Disaggregated InfrastructureQiaolin Yu, Chang Guo, Jay Zhuang, Viraj Thakkar 等SIGMOD 2024 · 被引用 18 次
- HiDPU: A DPU-Oriented Hybrid Indexing Scheme for Disaggregated Storage SystemsWenbin Zhu, Zhaoyan Shen, Qian Wei, Renhai Chen 等FAST 2025 · 被引用 11 次
- CXL Memory Performance for In-Memory Data ProcessingMarcel Weisgut, Daniel Ritter, Pinar Tözün, Lawrence Benson 等VLDB 2025 · 被引用 7 次
- Outback: Fast and Communication-efficient Index for Key-Value Store on Disaggregated MemoryYi Liu, Minghao Xie, Shouqian Shi, Yuanchao Xu 等VLDB 2025 · 被引用 5 次
- CloudyBench: A Testbed for A Comprehensive Evaluation of Cloud-Native DatabasesChao Zhang, Guoliang Li, Leyao Liu, Tao Lv 等ICDE 2025 · 被引用 4 次
它引用的顶会 Paper9
- Building An Elastic Query Engine on Disaggregated StorageMidhul Vuppalapati, Justin Miron, Rachit Agarwal, Dan Truong 等NSDI 2020 · 被引用 142 次
- Sherman: A Write-Optimized Distributed B+Tree Index on Disaggregated MemoryQing Wang, Youyou Lu, Jiwu ShuSIGMOD 2022 · 被引用 99 次
- FlexPushdownDB: Hybrid Pushdown and Caching in a Cloud DBMSYifei Yang, Matt Youill, Matthew E. Woicik, Yizhou Liu 等VLDB 2021 · 被引用 67 次
- Towards Cost-Effective and Elastic Cloud Database Deployment via Memory DisaggregationYingqiang Zhang, Chaoyi Ruan, Cheng Li, Jimmy Yang 等VLDB 2021 · 被引用 54 次
- Hailstorm: Disaggregated Compute and Storage for Distributed LSM-based DatabasesLaurent Bindschaedler, Ashvin Goel, Willy ZwaenepoelASPLOS 2020 · 被引用 51 次
相关 Paper
- Reducing Tail Latency in Storage-Disaggregated Database SystemsXi Pang, Jianguo WangSIGMOD 2026 · 被引用 2 次
- Understanding and Optimizing Database Pushdown on Disaggregated StorageHua Zhang, Xiao Li, Yuebin Bai, Ming LiuASPLOS 2026 · 被引用 1 次
- Understanding the Effect of Data Center Resource Disaggregation on Production DBMSsQizhen Zhang, Yifan Cai, Xinyi Chen, Sebastian Angel 等VLDB 2020 · 被引用 64 次
- Terark-DS: A High-Performance and Storage-Efficient Key-Value Separation Storage Engine on Disaggregated StorageJianshun Zhang, Xun Deng, Fang Wang, Jiaxin Ou 等VLDB 2026
- Marlin: Efficient Coordination for Autoscaling Cloud DBMSWenjie Hu, Guanzhou Hu, Mahesh Balakrishnan, Xiangyao YuSIGMOD 2026 · 被引用 1 次
