Virtual Consensus in Delos
Mahesh Balakrishnan, Jason Flinn, Chen Shen, Mihir Dharamshi, Ahmed Jafri, Xiao Shi, Santosh Ghosh, Hazem Hassan, Aaryaman Sagar, Rhed Shi, Jingming Liu, Filip Gruszczynski
Abstract
Consensus-based replicated systems are complex, monolithic, and difficult to upgrade once deployed. As a result, deployed systems do not benefit from innovative research, and new consensus protocols rarely reach production. We propose virtualizing consensus by virtualizing the shared log API, allowing services to change consensus protocols without downtime. Virtualization splits the logic of consensus into the VirtualLog, a generic and reusable reconfiguration layer; and pluggable ordering protocols called Loglets. Loglets are simple, since they do not need to support reconfiguration or leader election; diverse, consisting of different protocols, codebases, and even deployment modes; and composable, via RAID-like stacking and striping. We describe a production database called Delos 1 which leverages virtual consensus for rapid, incremental development and deployment. Delos reached production within 8 months, and 4 months later upgraded its consensus protocol without downtime for a 10X latency improvement. Delos can dynamically change its performance properties by changing consensus protocols: we can scale throughput by up to 10X by switching to a disaggregated Loglet, and double the failure threshold of an instance without sacrificing throughput via a striped Loglet.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 140d426a-f39c-4ed2-b35f-c4b7f64b2bdbCited by top-tier papers25
- Twine: A Unified Cluster Management System for Shared InfrastructureChunqiang Tang, Kenny Yu, Kaushik Veeraraghavan, Jonathan Kaldor et al.OSDI 2020 · 107 citations
- Boki: Stateful Serverless Computing with Shared LogsZhipeng Jia, Emmett WitchelSOSP 2021 · 81 citations
- Fault-Tolerant Replication with Pull-Based Consensus in MongoDBSiyuan Zhou, Shuai MuNSDI 2021 · 38 citations
- Halfmoon: Log-Optimal Fault-Tolerant Stateful Serverless ComputingSheng Qi, Xuanzhe Liu, Xin JinSOSP 2023 · 16 citations
- Fine-Grained Re-Execution for Efficient Batched Commit of Distributed TransactionsZhiyuan Dong, Zhaoguo Wang, Xiaodong Zhang, Xian Xu et al.VLDB 2023 · 16 citations
Builds on6
- Twine: A Unified Cluster Management System for Shared InfrastructureChunqiang Tang, Kenny Yu, Kaushik Veeraraghavan, Jonathan Kaldor et al.OSDI 2020 · 107 citations
- PigPaxos: Devouring the Communication Bottlenecks in Distributed ConsensusAleksey Charapko, Ailidani Ailijiang, Murat DemirbasSIGMOD 2021 · 55 citations
- HovercRaft: achieving scalability and fault-tolerance for microsecond-scale datacenter servicesMarios Kogias, Edouard BugnionEuroSys 2020 · 52 citations
- Scalog: Seamless Reconfiguration and Total Order in a Scalable Shared LogCong Ding, David Chu, Evan Zhao, Xiang Li et al.NSDI 2020 · 52 citations
- Gryff: Unifying Consensus and Shared RegistersMatthew Burke, Audrey Cheng, Wyatt LloydNSDI 2020 · 29 citations
Related papers
- Log-structured Protocols in DelosMahesh Balakrishnan, Chen Shen, Ahmed Jafri, Suyog Mapara et al.SOSP 2021 · 10 citations
- LeaseGuard: Raft Leases Done RightA. Jesse Jiryu Davis, Murat Demirbas, Lingzhi DengSIGMOD 2026 · 2 citations
- Jetpack: Consensus Made Generally FastZe Tang, Zihao Zhang, Weihai Shen, Jicheng Shi et al.OSDI 2026
- FLEET: High-Performance Durable Replicated State Machines using Scattered and Coordinated Log EntriesHua Fan, Hao Tan, Wenchao Zhou, Feifei LiVLDB 2025 · 1 citation
- Bodega: Localized Linearizable Reads at Anywhere Anytime via Roster LeasesGuanzhou Hu, Andrea C. Arpaci-Dusseau, Remzi H. Arpaci-DusseauOSDI 2026
