RIO: Order-Preserving and CPU-Efficient Remote Storage Access
Xiaojian Liao, Zhe Yang, Jiwu Shu
Abstract
Modern NVMe SSDs and RDMA networks provide dramatically higher bandwidth and concurrency. Existing networked storage systems (e.g., NVMe over Fabrics) fail to fully exploit these new devices due to inefficient storage ordering guarantees. Severe synchronous execution for storage order in these systems stalls the CPU and I/O devices and lowers the CPU and I/O performance efficiency of the storage system.
We present Rio, a new approach to the storage order of remote storage access. The key insight in Rio is that the layered design of the software stack, along with the concurrent and asynchronous network and storage devices, makes the storage stack conceptually similar to the CPU pipeline. Inspired by the CPU pipeline that executes out-of-order and commits in-order, Rio introduces the I/O pipeline that allows internal out-of-order and asynchronous execution for ordered write requests while offering intact external storage order to applications. Together with merging consecutive ordered requests, these design decisions make for write throughput and CPU efficiency close to that of orderless requests.
We implement Rio in Linux NVMe over RDMA stack, and further build a file system named RioFS atop Rio. Evaluations show that Rio outperforms Linux NVMe over RDMA and a state-of-the-art storage stack named Horae by two orders of magnitude and 4.9× on average in terms of throughput of ordered write requests, respectively. RioFS increases the throughput of RocksDB by 1.9× and 1.5× on average, against Ext4 and HoraeFS, respectively.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1446d8e9-d726-43fe-9df1-a7a2b8b06919Cited by top-tier papers2
- PIFS-Rec: Process-In-Fabric-Switch for Large-Scale Recommendation System InferencesPingyi Huo, Anusha Devulapally, Hasan Al Maruf, Minseo Park et al.MICRO 2024 · 6 citations
- CETOFS: A High-Performance File System with Host-Server Collaboration for Remote StorageWenqing Jia, Dejun Jiang, Jin XiongFAST 2026
Builds on8
- TCP ≈ RDMA: CPU-efficient Remote Storage Access with i10Jaehyun Hwang, Qizhe Cai, Ao Tang, Rachit AgarwalNSDI 2020 · 70 citations
- Characterizing and Optimizing Remote Persistent Memory with RDMA and NVMXingda Wei, Xiating Xie, Rong Chen, Haibo Chen et al.USENIX ATC 2021 · 48 citations
- Gimbal: enabling multi-tenant storage disaggregation on SmartNIC JBOFsJaehong Min, Ming Liu, Tapan Chugh, Chenxingyu Zhao et al.SIGCOMM 2021 · 47 citations
- Max: A Multicore-Accelerated File System for Flash StorageXiaojian Liao, Youyou Lu, Erci Xu, Jiwu ShuUSENIX ATC 2021 · 38 citations
- Write Dependency Disentanglement with HORAEXiaojian Liao, Youyou Lu, Erci Xu, Jiwu ShuOSDI 2020 · 31 citations
Related papers
- TeRM: Extending RDMA-Attached Memory with SSDZhe Yang, Qing Wang, Xiaojian Liao, Youyou Lu et al.FAST 2024 · 6 citations
- Exploring the Asynchrony of Slow Memory Filesystem with EasyIOBohong Zhu, Youmin Chen, Jiwu ShuEuroSys 2024 · 4 citations
- Hitchhike: Efficient Request Submission via Deferred Enforcement of Address ContiguityXuda Zheng, Jian Zhou, Shuhan Bai, Runjin Wu et al.ASPLOS 2026
- Direct File Path Access with NVMe over Fabrics for High-Performance Remote Data ReadsRohit Verma, Arun RaghunathINFOCOM 2026
- λ-IO: A Unified IO Stack for Computational StorageZhe Yang, Youyou Lu, Xiaojian Liao, Youmin Chen et al.FAST 2023 · 54 citations
