PigPaxos: Devouring the Communication Bottlenecks in Distributed Consensus
Aleksey Charapko, Ailidani Ailijiang, Murat Demirbas
摘要
Strong consistency replication helps keep application logic simple and provides significant benefits for correctness and manageability. Unfortunately, the adoption of strongly-consistent replication protocols has been curbed due to their limited scalability and performance. To alleviate the leader bottleneck in strongly-consistent replication protocols, we introduce Pig, an in-protocol communication aggregation and piggybacking technique. Pig employs randomly selected nodes from follower subgroups to relay the leader's message to the rest of the followers in the subgroup, and to perform in-network aggregation of acknowledgments back from these followers. By randomly alternating the relay nodes across replication operations, Pig shields the relay nodes as well as the leader from becoming hotspots and improves throughput scalability.
We showcase Pig in the context of classical Paxos protocols employed for strongly consistent replication by many cloud computing services and databases. We implement and evaluate PigPaxos, in comparison to Paxos and EPaxos protocols under various workloads over clusters of size 5 to 25 nodes. We show that the aggregation at the relay has little latency overhead, and PigPaxos can provide more than 3 folds improved throughput over Paxos and EPaxos with little latency deterioration. We support our experimental observations with the analytical modeling of the bottlenecks and show that the rotating of the relay nodes provides the most benefit for reducing the bottlenecks and that the throughput is maximized when employing only 1 randomly rotating relay node.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper19
- Virtual Consensus in DelosMahesh Balakrishnan, Jason Flinn, Chen Shen, Mihir Dharamshi 等OSDI 2020 · 被引用 42 次
- Scaling Replicated State Machines with CompartmentalizationMichael J. Whittaker, Ailidani Ailijiang, Aleksey Charapko, Murat Demirbas 等VLDB 2021 · 被引用 40 次
- The Bedrock of Byzantine Fault Tolerance: A Unified Platform for BFT Protocols Analysis, Implementation, and ExperimentationMohammad Javad Amiri, Chenyuan Wu, Divyakant Agrawal, Amr El Abbadi 等NSDI 2024 · 被引用 39 次
- Fault-Tolerant Replication with Pull-Based Consensus in MongoDBSiyuan Zhou, Shuai MuNSDI 2021 · 被引用 38 次
- Scaling Blockchain Consensus via a Robust Shared MempoolFangyu Gai, Jianyu Niu, Ivan Beschastnikh, Chen Feng 等ICDE 2023 · 被引用 25 次
它引用的顶会 Paper2
- Enhancing Bitcoin Security and Performance with Strong Consistency via Collective SigningEleftherios Kokoris-Kogias, Philipp Jovanovic, Nicolas Gailly, Ismail Khoffi 等USENIX Security 2016 · 被引用 769 次
- ResilientDB: Global Scale Resilient Blockchain FabricSuyash Gupta, Sajjad Rahnama, Jelle Hellings, Mohammad SadoghiVLDB 2020 · 被引用 100 次
相关 Paper
- RL-Paxos: Relieving the Leader's Burden with Efficient Task Offloading in Distributed ConsensusChenhao Zhang, Jinquan Wang, Meng Han, Bing Wei 等ICDE 2026
- SwiftPaxos: Fast Geo-Replicated State MachinesFedor Ryabinin, Alexey Gotsman, Pierre SutraNSDI 2024 · 被引用 19 次
- State machine replication scalability made simpleChrysoula Stathakopoulou, Matej Pavlovic, Marko VukolicEuroSys 2022 · 被引用 55 次
- FLAIR: Accelerating Reads with Consistency-Aware Network RoutingHatem Takruri, Ibrahim Kettaneh, Ahmed Alquraan, Samer Al-KiswanyNSDI 2020 · 被引用 20 次
- Efficient replication via timestamp stabilityVitor Enes, Carlos Baquero, Alexey Gotsman, Pierre SutraEuroSys 2021 · 被引用 21 次
