HadaFS: A File System Bridging the Local and Shared Burst Buffer for Exascale Supercomputers
Xiaobin He, Bin Yang, Jie Gao, Wei Xiao, Qi Chen, Shupeng Shi, Dexun Chen, Weiguo Liu, Wei Xue, Zuoning Chen
摘要
Current supercomputers introduce SSDs to form a Burst Buffer (BB) layer to meet the HPC application's growing I/O requirements. BBs can be divided into two types by deployment location. One is the local BB, which is known for its scalability and performance. The other is the shared BB, which has the advantage of data sharing and deployment costs. How to unify the advantages of the local BB and the shared BB is a key issue in the HPC community.
We propose a novel BB file system named HadaFS that provides the advantages of local BB deployments to shared BB deployments. First, HadaFS offers a new Localized Triage Architecture (LTA) to solve the problem of ultra-scale expansion and data sharing. Then, HadaFS proposes a full-path indexing approach with three metadata synchronization strategies to solve the problem of complex metadata management of traditional file systems and mismatch with the application I/O behaviors. Moreover, HadaFS integrates a data management tool named Hadash, which supports efficient data query in the BB and accelerates data migration between the BB and traditional HPC storage. HadaFS has been deployed on the Sunway New-generation Supercomputer (SNS), serving hundreds of applications and supporting a maximum of 600,000-client scaling.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper2
- Access Patterns and Performance Behaviors of Multi-layer Supercomputer I/O Subsystems under Production LoadJean Luca Bez, Ahmad Maroof Karimi, Arnab Kumar Paul, Bing Xie 等HPDC 2022 · 被引用 26 次
- File System Semantics Requirements of HPC ApplicationsChen Wang, Kathryn Mohror, Marc SnirHPDC 2021 · 被引用 25 次
相关 Paper
- Fine-grained Policy-driven I/O Sharing for Burst BuffersEd Karrels, Lei Huang, Yuhong Kan, Ishank Arora 等SC 2023 · 被引用 5 次
- DeltaFS: a scalable no-ground-truth filesystem for massively-parallel computingQing Zheng, Charles D. Cranor, Gregory R. Ganger, Garth A. Gibson 等SC 2021 · 被引用 5 次
- Concealing Compression-accelerated I/O for HPC Applications through In Situ Task SchedulingSian Jin, Sheng Di, Frédéric Vivien, Daoce Wang 等EuroSys 2024 · 被引用 13 次
- HAVS: Hardware-accelerated Shared-memory-based VPP Network StackShujun Zhuang, Jian Zhao, Jian Li, Ping Yu 等INFOCOM 2021 · 被引用 1 次
- BAASH: lightweight, efficient, and reliable blockchain-as-a-service for HPC systemsAbdullah Al-Mamun, Feng Yan, Dongfang ZhaoSC 2021 · 被引用 14 次
