RIO: Order-Preserving and CPU-Efficient Remote Storage Access
Xiaojian Liao, Zhe Yang, Jiwu Shu
摘要
Modern NVMe SSDs and RDMA networks provide dramatically higher bandwidth and concurrency. Existing networked storage systems (e.g., NVMe over Fabrics) fail to fully exploit these new devices due to inefficient storage ordering guarantees. Severe synchronous execution for storage order in these systems stalls the CPU and I/O devices and lowers the CPU and I/O performance efficiency of the storage system.
We present Rio, a new approach to the storage order of remote storage access. The key insight in Rio is that the layered design of the software stack, along with the concurrent and asynchronous network and storage devices, makes the storage stack conceptually similar to the CPU pipeline. Inspired by the CPU pipeline that executes out-of-order and commits in-order, Rio introduces the I/O pipeline that allows internal out-of-order and asynchronous execution for ordered write requests while offering intact external storage order to applications. Together with merging consecutive ordered requests, these design decisions make for write throughput and CPU efficiency close to that of orderless requests.
We implement Rio in Linux NVMe over RDMA stack, and further build a file system named RioFS atop Rio. Evaluations show that Rio outperforms Linux NVMe over RDMA and a state-of-the-art storage stack named Horae by two orders of magnitude and 4.9× on average in terms of throughput of ordered write requests, respectively. RioFS increases the throughput of RocksDB by 1.9× and 1.5× on average, against Ext4 and HoraeFS, respectively.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- PIFS-Rec: Process-In-Fabric-Switch for Large-Scale Recommendation System InferencesPingyi Huo, Anusha Devulapally, Hasan Al Maruf, Minseo Park 等MICRO 2024 · 被引用 6 次
- CETOFS: A High-Performance File System with Host-Server Collaboration for Remote StorageWenqing Jia, Dejun Jiang, Jin XiongFAST 2026
它引用的顶会 Paper8
- TCP ≈ RDMA: CPU-efficient Remote Storage Access with i10Jaehyun Hwang, Qizhe Cai, Ao Tang, Rachit AgarwalNSDI 2020 · 被引用 70 次
- Characterizing and Optimizing Remote Persistent Memory with RDMA and NVMXingda Wei, Xiating Xie, Rong Chen, Haibo Chen 等USENIX ATC 2021 · 被引用 48 次
- Gimbal: enabling multi-tenant storage disaggregation on SmartNIC JBOFsJaehong Min, Ming Liu, Tapan Chugh, Chenxingyu Zhao 等SIGCOMM 2021 · 被引用 47 次
- Max: A Multicore-Accelerated File System for Flash StorageXiaojian Liao, Youyou Lu, Erci Xu, Jiwu ShuUSENIX ATC 2021 · 被引用 38 次
- Write Dependency Disentanglement with HORAEXiaojian Liao, Youyou Lu, Erci Xu, Jiwu ShuOSDI 2020 · 被引用 31 次
相关 Paper
- TeRM: Extending RDMA-Attached Memory with SSDZhe Yang, Qing Wang, Xiaojian Liao, Youyou Lu 等FAST 2024 · 被引用 6 次
- Exploring the Asynchrony of Slow Memory Filesystem with EasyIOBohong Zhu, Youmin Chen, Jiwu ShuEuroSys 2024 · 被引用 4 次
- Hitchhike: Efficient Request Submission via Deferred Enforcement of Address ContiguityXuda Zheng, Jian Zhou, Shuhan Bai, Runjin Wu 等ASPLOS 2026
- Direct File Path Access with NVMe over Fabrics for High-Performance Remote Data ReadsRohit Verma, Arun RaghunathINFOCOM 2026
- λ-IO: A Unified IO Stack for Computational StorageZhe Yang, Youyou Lu, Xiaojian Liao, Youmin Chen 等FAST 2023 · 被引用 54 次
