How to Copy Files
Yang Zhan, Alexander Conway, Yizheng Jiao, Nirjhar Mukherjee, Ian Groombridge, Michael A. Bender, Martin Farach-Colton, William Jannen, Rob Johnson, Donald E. Porter, Jun Yuan
摘要
Making logical copies, or clones, of files and directories is critical to many real-world applications and workflows, including backups, virtual machines, and containers. An ideal clone implementation meets the following performance goals: (1) creating the clone has low latency; (2) reads are fast in all versions (i.e., spatial locality is always maintained, even after modifications); (3) writes are fast in all versions; (4) the overall system is space efficient. Implementing a clone operation that realizes all four properties, which we call a nimble clone, is a long-standing open problem.
This paper describes nimble clones in BetrFS, an opensource, full-path-indexed, and write-optimized file system. The key observation behind our work is that standard copyon-write heuristics can be too coarse to be space efficient, or too fine-grained to preserve locality. On the other hand, a write-optimized key-value store, as used in BetrFS or an LSMtree, can decouple the logical application of updates from the granularity at which data is physically copied. In our writeoptimized clone implementation, data sharing among clones is only broken when a clone has changed enough to warrant making a copy, a policy we call copy-on-abundant-write.
We demonstrate that the algorithmic work needed to batch and amortize the cost of BetrFS clone operations does not erode the performance advantages of baseline BetrFS; BetrFS performance even improves in a few cases. BetrFS cloning is efficient; for example, when using the clone operation for container creation, BetrFS outperforms a simple recursive copy by up to two orders-of-magnitude and outperforms file systems that have specialized LXC backends by 3-4×.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- RunD: A Lightweight Secure Container Runtime for High-density Deployment and High-concurrency Startup in Serverless ComputingZijun Li, Jiagan Cheng, Quan Chen, Eryu Guan 等USENIX ATC 2022 · 被引用 106 次
- SplinterDB: Closing the Bandwidth Gap for NVMe Key-Value StoresAlexander Conway, Abhishek Gupta, Vijay Chidambaram, Martin Farach-Colton 等USENIX ATC 2020 · 被引用 90 次
- Remap-SSD: Safely and Efficiently Exploiting SSD Address Remapping to Eliminate Duplicate WritesYou Zhou, Qiulin Wu, Fei Wu, Hong Jiang 等FAST 2021 · 被引用 39 次
- BetrFS: a compleat file system for commodity SSDsYizheng Jiao, Simon Bertron, Sagar Patel, Luke Zeller 等EuroSys 2022 · 被引用 4 次
- SolFS: An Operation-Log Versioning File System for Hash-free Efficient Mobile Cloud BackupRiwei Pan, Yu Liang, Lei Li, Hongchao Du 等USENIX ATC 2025
相关 Paper
- NetClone: Fast, Scalable, and Dynamic Request Cloning for Microsecond-Scale RPCsGyuyeong KimSIGCOMM 2023 · 被引用 4 次
- MetaWBC: POSIX-Compliant Metadata Write-Back Caching for Distributed File SystemsYingjin Qian, Wen Cheng, Lingfang Zeng, Marc-André Vef 等SC 2022 · 被引用 6 次
- RubbleDB: CPU-Efficient Replication with NVMe-oFHaoyu Li, Sheng Jiang, Chen Chen, Ashwini Raina 等USENIX ATC 2023 · 被引用 9 次
- NobLSM: an LSM-tree with non-blocking writes for SSDsHaoran Dang, Chongnan Ye, Yanpeng Hu, Chundong WangDAC 2022 · 被引用 5 次
- Enabling High-Performance and Secure Userspace NVM File Systems with the Trio ArchitectureDiyu Zhou, Vojtech Aschenbrenner, Tao Lyu, Jian Zhang 等SOSP 2023 · 被引用 9 次
