State-machine replication for planet-scale systems
Vitor Enes, Carlos Baquero, Tuanir França Rezende, Alexey Gotsman, Matthieu Perrin, Pierre Sutra
Abstract
Online applications now routinely replicate their data at multiple sites around the world. In this paper we present Atlas, the first state-machine replication protocol tailored for such planet-scale systems. Atlas does not rely on a distinguished leader, so clients enjoy the same quality of service independently of their geographical locations. Furthermore, clientperceived latency improves as we add sites closer to clients. To achieve this, Atlas minimizes the size of its quorums using an observation that concurrent data center failures are rare. It also processes a high percentage of accesses in a single round trip, even when these conflict. We experimentally demonstrate that Atlas consistently outperforms state-of-the-art protocols in planet-scale scenarios. In particular, Atlas is up to two times faster than Flexible Paxos with identical failure assumptions, and more than doubles the performance of Egalitarian Paxos in the YCSB benchmark.
• Theory of computation → Distributed algorithms.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext afe462e8-5a0a-441e-813a-905958fd3919Cited by top-tier papers16
- Efficient replication via timestamp stabilityVitor Enes, Carlos Baquero, Alexey Gotsman, Pierre SutraEuroSys 2021 · 21 citations
- Rabia: Simplifying State-Machine Replication Through RandomizationHaochen Pan, Jesse Tuglu, Neo Zhou, Tianshu Wang et al.SOSP 2021 · 20 citations
- SwiftPaxos: Fast Geo-Replicated State MachinesFedor Ryabinin, Alexey Gotsman, Pierre SutraNSDI 2024 · 19 citations
- Odyssey: the impact of modern hardware on strongly-consistent replication protocolsVasilis Gavrielatos, Antonios Katsarakis, Vijay NagarajanEuroSys 2021 · 11 citations
- QuePaxa: Escaping the tyranny of timeouts in consensusPasindu Tennage, Cristina Basescu, Lefteris Kokoris-Kogias, Ewa Syta et al.SOSP 2023 · 7 citations
Builds on1
Related papers
- RL-Paxos: Relieving the Leader's Burden with Efficient Task Offloading in Distributed ConsensusChenhao Zhang, Jinquan Wang, Meng Han, Bing Wei et al.ICDE 2026
- HovercRaft: achieving scalability and fault-tolerance for microsecond-scale datacenter servicesMarios Kogias, Edouard BugnionEuroSys 2020 · 52 citations
- Tolerating Slowdowns in Replicated State Machines using CopilotsKhiem Ngo, Siddhartha Sen, Wyatt LloydOSDI 2020 · 24 citations
- Omni-Paxos: Breaking the Barriers of Partial ConnectivityHarald Ng, Seif Haridi, Paris CarboneEuroSys 2023 · 8 citations
- Bandle: Asynchronous State Machine Replication Made EfficientBo Wang, Shengyun Liu, He Dong, Xiangzhe Wang et al.EuroSys 2024 · 4 citations
