Rearchitecting Linux Storage Stack for µs Latency and High Throughput
Jaehyun Hwang, Midhul Vuppalapati, Simon Peter, Rachit Agarwal
摘要
This paper demonstrates that it is possible to achieve µs-scale latency using Linux kernel storage stack, even when tens of latency-sensitive applications compete for host resources with throughput-bound applications that perform read/write operations at throughput close to hardware capacity. Furthermore, such performance can be achieved without any modification in applications, network hardware, kernel CPU schedulers and/or kernel network stack.
We demonstrate the above using design, implementation and evaluation of blk-switch, a new Linux kernel storage stack architecture. The key insight in blk-switch is that Linux's multi-queue storage design, along with multi-queue network and storage hardware, makes the storage stack conceptually similar to a network switch. blk-switch uses this insight to adapt techniques from the computer networking literature (e.g., multiple egress queues, prioritized processing of individual requests, load balancing, and switch scheduling) to the Linux kernel storage stack.
blk-switch evaluation over a variety of scenarios shows that it consistently achieves µs-scale average and tail latency (at both 99 th and 99.9 th percentiles), while allowing applications to near-perfectly utilize the hardware capacity.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper25
- Empowering Azure Storage with RDMAWei Bai, Shanim Sainul Abdeen, Ankit Agrawal, Krishan Kumar Attre 等NSDI 2023 · 被引用 117 次
- XRP: In-Kernel Storage Functions with eBPFYuhong Zhong, Haoyu Li, Yu Jian Wu, Ioannis Zarkadas 等OSDI 2022 · 被引用 100 次
- Efficient Scheduling Policies for Microsecond-Scale TasksSarah McClure, Amy Ousterhout, Scott Shenker, Sylvia RatnasamyNSDI 2022 · 被引用 43 次
- ODINFS: Scaling PM Performance with Opportunistic DelegationDiyu Zhou, Yuchen Qian, Vishal Gupta, Zhifei Yang 等OSDI 2022 · 被引用 29 次
- Karma: Resource Allocation for Dynamic DemandsMidhul Vuppalapati, Giannis Fikioris, Rachit Agarwal, Asaf Cidon 等OSDI 2023 · 被引用 22 次
它引用的顶会 Paper3
- Caladan: Mitigating Interference at Microsecond TimescalesJoshua Fried, Zhenyuan Ruan, Amy Ousterhout, Adam BelayOSDI 2020 · 被引用 213 次
- Building An Elastic Query Engine on Disaggregated StorageMidhul Vuppalapati, Justin Miron, Rachit Agarwal, Dan Truong 等NSDI 2020 · 被引用 142 次
- TCP ≈ RDMA: CPU-efficient Remote Storage Access with i10Jaehyun Hwang, Qizhe Cai, Ao Tang, Rachit AgarwalNSDI 2020 · 被引用 70 次
相关 Paper
- SKQ: Event Scheduling for Optimizing Tail Latency in a Traditional OS KernelSiyao Zhao, Haoyu Gu, Ali José MashtizadehUSENIX ATC 2021 · 被引用 11 次
- Towards μs tail latency and terabit ethernet: disaggregating the host network stackQizhe Cai, Midhul Vuppalapati, Jaehyun Hwang, Christos Kozyrakis 等SIGCOMM 2022 · 被引用 20 次
- Opening Up Kernel-Bypass TCP StacksShinichi Awamoto, Michio HondaUSENIX ATC 2025 · 被引用 5 次
- Understanding Host Network Stack LatencyTianyu Zuo, Jaehyun Hwang, Ao Tang, Rachit Agarwal 等SIGCOMM 2026
- λ-IO: A Unified IO Stack for Computational StorageZhe Yang, Youyou Lu, Xiaojian Liao, Youmin Chen 等FAST 2023 · 被引用 54 次
