CFS: Scaling Metadata Service for Distributed File System via Pruned Scope of Critical Sections
Yiduo Wang, Yufei Wu, Cheng Li, Pengfei Zheng, Biao Cao, Yan Sun, Fei Zhou, Yinlong Xu, Yao Wang, Guangjun Xie
摘要
There is a fundamental tension between metadata scalability and POSIX semantics within distributed file systems. The bottleneck lies in the coordination, mainly locking, used for ensuring strong metadata consistency, namely, atomicity and isolation. CFS is a scalable, fully POSIX-compliant distributed file system that eliminates the metadata management bottleneck via pruning the scope of critical sections for reduced locking overhead. First, CFS adopts a tiered metadata organization to scale file attributes and the remaining namespace hierarchies independently with appropriate partitioning and indexing methods, eliminating cross-shard distributed coordination. Second, it further scales up the single metadata shard performance by single-shard atomic primitives, shortening the metadata requests' lifespan and removing spurious conflicts. Third, CFS drops the metadata proxy layer but employs the light-weight, scalable client-side metadata resolving. CFS has been running in the production environment of Baidu AI Cloud for three years. Our evaluation with a 50-node cluster and microbenchmarks shows that CFS simultaneously improves the throughput of baselines like HopsFS and InfiniFS by 1.76--75.82× and 1.22--4.10×, and reduces their average latency by up to 91.71% and 54.54%, respectively. Under cases with higher contention and larger directories, CFS' throughput benefits expand by one order of magnitude. For three real-world workloads with data accesses, CFS introduces 1.62--2.55× end-to-end throughput speedups and 35.06--62.47% tail latency reductions over InfiniFS.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper4
- SingularFS: A Billion-Scale Distributed File System Using a Single Metadata ServerHao Guo, Youyou Lu, Wenhao Lv, Xiaojian Liao 等USENIX ATC 2023 · 被引用 17 次
- FalconFS: Distributed File System for Large-Scale Deep Learning PipelineJingwei Xu, Junbin Kang, Mingkai Dong, Mingyu Liu 等NSDI 2026 · 被引用 2 次
- SwitchFS: Asynchronous Metadata Updates for Distributed Filesystems with In-Network CoordinationJingwei Xu, Mingkai Dong, Qiulin Tian, Ziyi Tian 等EuroSys 2026 · 被引用 1 次
- Accelerating Metadata Management of DFS via Speculative Permission CheckingYiduo Wang, Linghang Meng, Liang Li, Jie WuICDE 2026
相关 Paper
- InfiniFS: An Efficient Metadata Service for Large-Scale Distributed FilesystemsWenhao Lv, Youyou Lu, Yiming Zhang, Peile Duan 等FAST 2022 · 被引用 52 次
- DeltaFS: a scalable no-ground-truth filesystem for massively-parallel computingQing Zheng, Charles D. Cranor, Gregory R. Ganger, Garth A. Gibson 等SC 2021 · 被引用 5 次
- λFS: A Scalable and Elastic Distributed File System Metadata Service using Serverless FunctionsBenjamin Carver, Runzhou Han, Jingyuan Zhang, Mai Zheng 等ASPLOS 2023 · 被引用 3 次
- Scalable Persistent Memory File System with Kernel-Userspace CollaborationYoumin Chen, Youyou Lu, Bohong Zhu, Andrea C. Arpaci-Dusseau 等FAST 2021 · 被引用 15 次
- DeLFS: A Decentralized Log-Structured File System for ManycoresTaehwan Ahn, Chanhyeong Yu, Sangjin Lee, Yongseok SonOSDI 2026
