In-Network Leaderless Replication for Distributed Data Stores
Gyuyeong Kim, Wonjun Lee
摘要
Leaderless replication allows any replica to handle any type of request to achieve read scalability and high availability for distributed data stores. However, this entails burdensome coordination overhead of replication protocols, degrading write throughput. In addition, the data store still requires coordination for membership changes, making it hard to resolve server failures quickly. To this end, we present NetLR, a replicated data store architecture that supports high performance, fault tolerance, and linearizability simultaneously. The key idea of NetLR is moving the entire replication functions into the network by leveraging the switch as an on-path in-network replication orchestrator. Specifically, NetLR performs consistency-aware read scheduling, high-performance write coordination, and active fault adaptation in the network switch. Our in-network replication eliminates inter-replica coordination for writes and membership changes, providing high write performance and fast failure handling. NetLR can be implemented using programmable switches at a line rate with only 5.68% of additional memory usage. We implement a prototype of NetLR on an Intel Tofino switch and conduct extensive testbed experiments. Our evaluation results show that NetLR is the only solution that achieves high throughput and low latency and is robust to server failures.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Pushing the Limits of In-Network Caching for Key-Value StoresGyuyeong KimNSDI 2025 · 被引用 9 次
- NetClone: Fast, Scalable, and Dynamic Request Cloning for Microsecond-Scale RPCsGyuyeong KimSIGCOMM 2023 · 被引用 4 次
- The LAW theorem: Local Reads and Linearizable Asynchronous ReplicationEmmanouil Giortamis, Antonios Katsarakis, Vasilis Gavrielatos, Pramod Bhatotia 等VLDB 2025 · 被引用 2 次
- SwitchFS: Asynchronous Metadata Updates for Distributed Filesystems with In-Network CoordinationJingwei Xu, Mingkai Dong, Qiulin Tian, Ziyi Tian 等EuroSys 2026 · 被引用 1 次
- Switch: Asynchronous Metadata Updating for Distributed Storage with in-Network Data VisibilityJunru Li, Qing Wang, Zhe Yang, Shuo Liu 等ICDE 2026
它引用的顶会 Paper5
- A large scale analysis of hundreds of in-memory cache clusters at TwitterJuncheng Yang, Yao Yue, K. V. RashmiOSDI 2020 · 被引用 245 次
- Pegasus: Tolerating Skewed Workloads in Distributed Storage with In-Network Coherence DirectoriesJialin Li, Jacob Nelson, Ellis Michael, Xin Jin 等OSDI 2020 · 被引用 96 次
- Harmonia: Near-Linear Scalability for Replicated Storage with In-Network Conflict DetectionHang Zhu, Zhihao Bai, Jialin Li, Ellis Michael 等VLDB 2020 · 被引用 58 次
- Hermes: A Fast, Fault-Tolerant and Linearizable Replication ProtocolAntonios Katsarakis, Vasilis Gavrielatos, M. R. Siavash Katebzadeh, Arpit Joshi 等ASPLOS 2020 · 被引用 47 次
- Characterizing, Modeling, and Benchmarking RocksDB Key-Value Workloads at FacebookZhichao Cao, Siying Dong, Sagar Vemuri, David H. C. DuFAST 2020
相关 Paper
- FLAIR: Accelerating Reads with Consistency-Aware Network RoutingHatem Takruri, Ibrahim Kettaneh, Ahmed Alquraan, Samer Al-KiswanyNSDI 2020 · 被引用 20 次
- In-Memory Key-Value Store Live Migration with NetMigrateZeying Zhu, Yibo Zhao, Zaoxing LiuFAST 2024 · 被引用 14 次
- RedPlane: enabling fault-tolerant stateful in-switch applicationsDaehyeok Kim, Jacob Nelson, Dan R. K. Ports, Vyas Sekar 等SIGCOMM 2021 · 被引用 33 次
- SwitchTx: Scalable In-Network Coordination for Distributed Transaction ProcessingJunru Li, Youyou Lu, Yiming Zhang, Qing Wang 等VLDB 2022 · 被引用 11 次
- P4KVS: A Role-Replica Separation Offloading Method to Achieve In-Network Consistency for KV Stores Based on P4 SwitchesHaojuan Li, Zongpu Zhang, Chenzhen Ye, Ruohan Tang 等SIGMOD 2026
