D2FQ: Device-Direct Fair Queueing for NVMe SSDs
Jiwon Woo, Minwoo Ahn, Gyusun Lee, Jinkyu Jeong
Abstract
With modern high-performance SSDs that can handle parallel I/O requests from multiple tenants, fair sharing of block I/O is an essential requirement for performance isolation. Typical block I/O schedulers take three steps (submit-arbitratedispatch) to transfer an I/O request to a device, and the three steps incur high overheads in terms of CPU utilization, scalability and block I/O performance. This motivates us to offload the I/O scheduling function to a device. If so, the three steps can be reduced to one step (submit=dispatch), thereby saving CPU cycles and improving the I/O performance.
To this end, we propose D2FQ, a fair-queueing I/O scheduler that exploits the NVMe weighted round-robin (WRR) arbitration, a device-side I/O scheduling feature. D2FQ abstracts the three classes of command queues in WRR as three queues with different I/O processing speeds. Then, for every I/O submission D2FQ selects and dispatches an I/O request to one of three queues immediately while satisfying fairness. This avoids time-consuming I/O scheduling operations, thereby saving CPU cycles and improving the block I/O performance. The prototype is implemented in the Linux kernel and evaluated with various workloads. With synthetic workloads, D2FQ provides fairness while saving CPU cycles by up to 45% as compared to MQFQ, a state-of-the-art fair queueing I/O scheduler.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c5e0e41f-e5f1-43cb-a885-2035c1ce1c92Cited by top-tier papers5
- LPNS: Scalable and Latency-Predictable Local Storage Virtualization for Unpredictable NVMe SSDs in CloudsBo Peng, Cheng Guo, Jianguo Yao, Haibing GuanUSENIX ATC 2023 · 10 citations
- LabStor: A Modular and Extensible Platform for Developing High-Performance, Customized I/O Stacks in UserspaceLuke Logan, Jaime Cernuda Garcia, Jay F. Lofstead, Xian-He Sun et al.SC 2022 · 8 citations
- Learning to Drive Software-Defined Solid-State DrivesDaixuan Li, Jinghan Sun, Jian HuangMICRO 2023 · 5 citations
- OPIMQ: Order Preserving IO stack for Multi-Queue Block DeviceJieun Kim, Joontaek Oh, Juwon Kim, Seung Won Yoo et al.FAST 2025 · 2 citations
- Espresso: Constructing Cost-Efficient CXL JBOF via Inter-SSD Computing Resource SharingShushu Yi, Yuda An, Li Peng, Xiurui Pan et al.OSDI 2026
Builds on1
Related papers
- Daredevil: Rescue Your Flash Storage from Inflexible Kernel Storage StackJunzhe Li, Ran Shu, Jiayi Lin, Qingyu Zhang et al.EuroSys 2025 · 2 citations
- Hitchhike: Efficient Request Submission via Deferred Enforcement of Address ContiguityXuda Zheng, Jian Zhou, Shuhan Bai, Runjin Wu et al.ASPLOS 2026
- Write Dependency Disentanglement with HORAEXiaojian Liao, Youyou Lu, Erci Xu, Jiwu ShuOSDI 2020 · 31 citations
- Fair Will Go On: A Collaboration-Aware Fairness Scheme for NVMe SSD in Cloud Storage SystemYang Zhou, Fang Wang, Zhan Shi, Dan Feng et al.DAC 2023 · 7 citations
- PipeSSD: A Lock-free Pipelined SSD Firmware Design for Multi-core ArchitectureZelin Du, Shaoqi Li, Zixuan Huang, Jin Xue et al.DAC 2024
