Morph: Efficient File-Lifetime Redundancy Management for Cluster File Systems
Timothy Kim, Sanjith Athlur, Saurabh Kadekodi, Francisco Maturana, Dax Delvira, Arif Merchant, Gregory R. Ganger, K. V. Rashmi
摘要
Many data services tune and change redundancy configurations of files over their lifetimes to address changes in data temperature and latency requirements. Unfortunately, changing redundancy configs (transcode) is IO-intensive. The Morph cluster file system introduces new transcode-efficient redundancy schemes to minimize overheads as files progress through lifetime phases. For newly ingested data, commonly stored via 3-way replication, Morph introduces a hybrid redundancy scheme that combines a replica with an erasurecoded (EC) stripe, reducing both ingest IO and capacity overheads while enabling free transcode to EC by deleting replicas. For subsequent transcodes to wider, more space-efficient EC configs, Morph exploits Convertible Codes, which minimize data read for EC transcode, and introduces new block placement policies to maximize their effectiveness.
Analysis of data ingest and transcode activity in Google storage clusters shows the current massive IO load and the potential savings from Morph's approach-transcode IO can be reduced by over 95%, and total ingest+transcode IO can be reduced by 50-60% while also reducing capacity overheads for newly ingested data by 20%. Experiments evaluating a Morph implementation in HDFS show that these benefits can be realized in a real system without hidden increases in complexity, tail latency, or degraded-mode latency.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Okapi: Decoupling Data Striping and Redundancy Grouping in Cluster File SystemsSanjith Athlur, Timothy Kim, Saurabh Kadekodi, Francisco Maturana 等OSDI 2025 · 被引用 1 次
- ACOS: Apple's Geo-Distributed Object Store at Exabyte ScaleBenjamin Baron, Aline Bousquet, Eric Metens, Swapnil Pimpale 等FAST 2026 · 被引用 1 次
- WiseCode: Breaking the Scalability Barriers of Wide-Stripe Vector CodesSijie Cai, Guangyan Zhang, Xiao NiuOSDI 2026
- TCO-driven Storage Provisioning for Exascale Data CentersTimothy Kim, Saurabh Kadekodi, Arif Merchant, Prashant Nema 等EuroSys 2026
- Scaling the IO Wall with Declarative IOSanjith Athlur, Sara McAllister, Theo Gregersen, Timothy Kim 等OSDI 2026
它引用的顶会 Paper5
- Facebook's Tectonic Filesystem: Efficiency from ExascaleSatadru Pan, Theano Stavrinos, Yunqiao Zhang, Atul Sikaria 等FAST 2021 · 被引用 110 次
- Practical Design Considerations for Wide Locally Recoverable Codes (LRCs)Saurabh Kadekodi, Shashwat Silas, David Clausen, Arif MerchantFAST 2023 · 被引用 55 次
- PACEMAKER: Avoiding HeART attacks in storage clusters with disk-adaptive redundancySaurabh Kadekodi, Francisco Maturana, Suhas Jayaram Subramanya, Juncheng Yang 等OSDI 2020 · 被引用 29 次
- Tiger: Disk-Adaptive Redundancy Without Placement RestrictionsSaurabh Kadekodi, Francisco Maturana, Sanjith Athlur, Arif Merchant 等OSDI 2022 · 被引用 20 次
- Thesios: Synthesizing Accurate Counterfactual I/O Traces from I/O SamplesPhitchaya Mangpo Phothilimthana, Saurabh Kadekodi, Soroush Ghodrati, Selene Moon 等ASPLOS 2024 · 被引用 6 次
相关 Paper
- Optimal Data Placement for Stripe Merging in Locally Repairable CodesSi Wu, Qingpeng Du, Patrick P. C. Lee, Yongkun Li 等INFOCOM 2022 · 被引用 23 次
- Balancing Repair Bandwidth and Sub-Packetization in Erasure-Coded Storage via Elastic TransformationKaicheng Tang, Keyun Cheng, Helen H. W. Chan, Xiaolu Li 等INFOCOM 2023 · 被引用 14 次
- LEGOStore: A Linearizable Geo-Distributed Store Combining Replication and Erasure CodingHamidReza Zare, Viveck R. Cadambe, Bhuvan Urgaonkar, Nader Alfares 等VLDB 2022 · 被引用 10 次
- MorphoSys: Automatic Physical Design Metamorphosis for Distributed Database SystemsMichael Abebe, Brad Glasbergen, Khuzaima DaudjeeVLDB 2020 · 被引用 18 次
- An exabyte a day: throughput-oriented, large scale, managed data transfers with EffingoLadislav Pápay, Jan Pustelnik, Krzysztof Rzadca, Beata Strack 等SIGCOMM 2024 · 被引用 11 次
