Scalable Parallel Flash Firmware for Many-core Architectures
Jie Zhang, Miryeong Kwon, Michael M. Swift, Myoungsoo Jung
摘要
NVMe is designed to unshackle flash from a traditional storage bus by allowing hosts to employ many threads to achieve higher bandwidth. While NVMe enables users to fully exploit all levels of parallelism offered by modern SSDs, current firmware designs are not scalable and have difficulty in handling a large number of I/O requests in parallel due to its limited computation power and many hardware contentions.
We propose DeepFlash, a novel manycore-based storage platform that can process more than a million I/O requests in a second (1MIOPS) while hiding long latencies imposed by its internal flash media. Inspired by a parallel data analysis system, we design the firmware based on many-to-many threading model that can be scaled horizontally. The proposed DeepFlash can extract the maximum performance of the underlying flash memory complex by concurrently executing multiple firmware components across many cores within the device. To show its extreme parallel scalability, we implement DeepFlash on a many-core prototype processor that employs dozens of lightweight cores, analyze new challenges from parallel I/O processing and address the challenges by applying concurrency-aware optimizations. Our comprehensive evaluation reveals that DeepFlash can serve around 4.5 GB/s, while minimizing the CPU demand on microbenchmarks and real server workloads.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- SpanDB: A Fast, Cost-Effective LSM-tree Based KV Store on Hybrid StorageHao Chen, Chaoyi Ruan, Cheng Li, Xiaosong Ma 等FAST 2021 · 被引用 120 次
- ZNS+: Advanced Zoned Namespace Interface for Supporting In-Storage Zone CompactionKyuhwa Han, Hyunho Gwak, Dongkun Shin, Jooyoung HwangOSDI 2021 · 被引用 107 次
- Behemoth: A Flash-centric Training Accelerator for Extreme-scale DNNsShine Kim, Yunho Jin, Gina Sohn, Jonghyun Bae 等FAST 2021 · 被引用 44 次
- Rearchitecting the TCP Stack for I/O-Offloaded Content DeliveryTaehyun Kim, Deondre Martin Ng, Junzhi Gong, Youngjin Kwon 等NSDI 2023 · 被引用 42 次
- Hardware/Software Co-Programmable Framework for Computational SSDs to Accelerate Deep Learning Service on Large-Scale GraphsMiryeong Kwon, Donghyun Gouk, Sangwon Lee, Myoungsoo JungFAST 2022 · 被引用 32 次
相关 Paper
- PipeSSD: A Lock-free Pipelined SSD Firmware Design for Multi-core ArchitectureZelin Du, Shaoqi Li, Zixuan Huang, Jin Xue 等DAC 2024
- What Modern NVMe Storage Can Do, And How To Exploit It: High-Performance I/O for High-Performance Storage EnginesGabriel Haas, Viktor LeisVLDB 2023 · 被引用 83 次
- BypassD: Enabling fast userspace access to shared SSDsSujay Yadalam, Chloe Alverti, Vasileios Karakostas, Jayneel Gandhi 等ASPLOS 2024 · 被引用 5 次
- Daredevil: Rescue Your Flash Storage from Inflexible Kernel Storage StackJunzhe Li, Ran Shu, Jiayi Lin, Qingyu Zhang 等EuroSys 2025 · 被引用 2 次
- Optimizing Memory-mapped I/O for Fast Storage DevicesAnastasios Papagiannis, Giorgos Xanthakis, Giorgos Saloustros, Manolis Marazakis 等USENIX ATC 2020 · 被引用 68 次
