Pegasus: Tolerating Skewed Workloads in Distributed Storage with In-Network Coherence Directories
Jialin Li, Jacob Nelson, Ellis Michael, Xin Jin, Dan R. K. Ports
摘要
High performance distributed storage systems face the challenge of load imbalance caused by skewed and dynamic workloads. This paper introduces Pegasus, a new storage system that leverages new-generation programmable switch ASICs to balance load across storage servers. Pegasus uses selective replication of the most popular objects in the data store to distribute load. Using a novel in-network coherence directory, the Pegasus switch tracks and manages the location of replicated objects. This allows it to achieve load-aware forwarding and dynamic rebalancing for replicated keys, while still guaranteeing data coherence and consistency. The Pegasus design is practical to implement as it stores only forwarding metadata in the switch data plane. The resulting system improves the throughput of a distributed in-memory key-value store by more than 10x under a latency SLO — results which hold across a large set of workloads with varying degrees of skew, read/write ratio, object sizes, and dynamism.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper33
- dLoRA: Dynamically Orchestrating Requests and Adapters for LoRA LLM ServingBingyang Wu, Ruidong Zhu, Zili Zhang, Peng Sun 等OSDI 2024 · 被引用 79 次
- Concordia: Distributed Shared Memory with In-Network Cache CoherenceQing Wang, Youyou Lu, Erci Xu, Junru Li 等FAST 2021 · 被引用 74 次
- MIND: In-Network Memory Management for Disaggregated Data CentersSeung-Seob Lee, Yanpeng Yu, Yupeng Tang, Anurag Khandelwal 等SOSP 2021 · 被引用 51 次
- Bidl: A High-throughput, Low-latency Permissioned Blockchain Framework for Datacenter NetworksJi Qi, Xusheng Chen, Yunpeng Jiang, Jianyu Jiang 等SOSP 2021 · 被引用 36 次
- RedPlane: enabling fault-tolerant stateful in-switch applicationsDaehyeok Kim, Jacob Nelson, Dan R. K. Ports, Vyas Sekar 等SIGCOMM 2021 · 被引用 33 次
它引用的顶会 Paper5
- A large scale analysis of hundreds of in-memory cache clusters at TwitterJuncheng Yang, Yao Yue, K. V. RashmiOSDI 2020 · 被引用 245 次
- The CacheLib Caching Engine: Design and Experiences at ScaleBenjamin Berg, Daniel S. Berger, Sara McAllister, Isaac Grosof 等OSDI 2020 · 被引用 145 次
- HotRing: A Hotspot-Aware In-Memory Key-Value StoreJiqiang Chen, Liang Chen, Sheng Wang, Guoyun Zhu 等FAST 2020 · 被引用 80 次
- Harmonia: Near-Linear Scalability for Replicated Storage with In-Network Conflict DetectionHang Zhu, Zhihao Bai, Jialin Li, Ellis Michael 等VLDB 2020 · 被引用 58 次
- Prism: Proxies without the PainYutaro Hayakawa, Michio Honda, Douglas Santry, Lars EggertNSDI 2021 · 被引用 32 次
相关 Paper
- Switch: Asynchronous Metadata Updating for Distributed Storage with in-Network Data VisibilityJunru Li, Qing Wang, Zhe Yang, Shuo Liu 等ICDE 2026
- SwitchFS: Asynchronous Metadata Updates for Distributed Filesystems with In-Network CoordinationJingwei Xu, Mingkai Dong, Qiulin Tian, Ziyi Tian 等EuroSys 2026 · 被引用 1 次
- FarReach: Write-back Caching in Programmable SwitchesSiyuan Sheng, Huancheng Puyang, Qun Huang, Lu Tang 等USENIX ATC 2023 · 被引用 22 次
- FLAIR: Accelerating Reads with Consistency-Aware Network RoutingHatem Takruri, Ibrahim Kettaneh, Ahmed Alquraan, Samer Al-KiswanyNSDI 2020 · 被引用 20 次
- P4DB - The Case for In-Network OLTPMatthias Jasny, Lasse Thostrup, Tobias Ziegler, Carsten BinnigSIGMOD 2022 · 被引用 18 次
