Virtual Consensus in Delos
Mahesh Balakrishnan, Jason Flinn, Chen Shen, Mihir Dharamshi, Ahmed Jafri, Xiao Shi, Santosh Ghosh, Hazem Hassan, Aaryaman Sagar, Rhed Shi, Jingming Liu, Filip Gruszczynski
摘要
Consensus-based replicated systems are complex, monolithic, and difficult to upgrade once deployed. As a result, deployed systems do not benefit from innovative research, and new consensus protocols rarely reach production. We propose virtualizing consensus by virtualizing the shared log API, allowing services to change consensus protocols without downtime. Virtualization splits the logic of consensus into the VirtualLog, a generic and reusable reconfiguration layer; and pluggable ordering protocols called Loglets. Loglets are simple, since they do not need to support reconfiguration or leader election; diverse, consisting of different protocols, codebases, and even deployment modes; and composable, via RAID-like stacking and striping. We describe a production database called Delos 1 which leverages virtual consensus for rapid, incremental development and deployment. Delos reached production within 8 months, and 4 months later upgraded its consensus protocol without downtime for a 10X latency improvement. Delos can dynamically change its performance properties by changing consensus protocols: we can scale throughput by up to 10X by switching to a disaggregated Loglet, and double the failure threshold of an instance without sacrificing throughput via a striped Loglet.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper25
- Twine: A Unified Cluster Management System for Shared InfrastructureChunqiang Tang, Kenny Yu, Kaushik Veeraraghavan, Jonathan Kaldor 等OSDI 2020 · 被引用 107 次
- Boki: Stateful Serverless Computing with Shared LogsZhipeng Jia, Emmett WitchelSOSP 2021 · 被引用 81 次
- Fault-Tolerant Replication with Pull-Based Consensus in MongoDBSiyuan Zhou, Shuai MuNSDI 2021 · 被引用 38 次
- Halfmoon: Log-Optimal Fault-Tolerant Stateful Serverless ComputingSheng Qi, Xuanzhe Liu, Xin JinSOSP 2023 · 被引用 16 次
- Fine-Grained Re-Execution for Efficient Batched Commit of Distributed TransactionsZhiyuan Dong, Zhaoguo Wang, Xiaodong Zhang, Xian Xu 等VLDB 2023 · 被引用 16 次
它引用的顶会 Paper6
- Twine: A Unified Cluster Management System for Shared InfrastructureChunqiang Tang, Kenny Yu, Kaushik Veeraraghavan, Jonathan Kaldor 等OSDI 2020 · 被引用 107 次
- PigPaxos: Devouring the Communication Bottlenecks in Distributed ConsensusAleksey Charapko, Ailidani Ailijiang, Murat DemirbasSIGMOD 2021 · 被引用 55 次
- HovercRaft: achieving scalability and fault-tolerance for microsecond-scale datacenter servicesMarios Kogias, Edouard BugnionEuroSys 2020 · 被引用 52 次
- Scalog: Seamless Reconfiguration and Total Order in a Scalable Shared LogCong Ding, David Chu, Evan Zhao, Xiang Li 等NSDI 2020 · 被引用 52 次
- Gryff: Unifying Consensus and Shared RegistersMatthew Burke, Audrey Cheng, Wyatt LloydNSDI 2020 · 被引用 29 次
相关 Paper
- Log-structured Protocols in DelosMahesh Balakrishnan, Chen Shen, Ahmed Jafri, Suyog Mapara 等SOSP 2021 · 被引用 10 次
- LeaseGuard: Raft Leases Done RightA. Jesse Jiryu Davis, Murat Demirbas, Lingzhi DengSIGMOD 2026 · 被引用 2 次
- Jetpack: Consensus Made Generally FastZe Tang, Zihao Zhang, Weihai Shen, Jicheng Shi 等OSDI 2026
- FLEET: High-Performance Durable Replicated State Machines using Scattered and Coordinated Log EntriesHua Fan, Hao Tan, Wenchao Zhou, Feifei LiVLDB 2025 · 被引用 1 次
- Bodega: Localized Linearizable Reads at Anywhere Anytime via Roster LeasesGuanzhou Hu, Andrea C. Arpaci-Dusseau, Remzi H. Arpaci-DusseauOSDI 2026
