USENIX ATC2023顶会
Light-Dedup: A Light-weight Inline Deduplication Framework for Non-Volatile Memory File Systems
Jiansheng Qiu, Yanqi Pan, Wen Xia, Xiaojia Huang, Wenjun Wu, Xiangyu Zou, Shiyi Li, Yu Hua
摘要
Emerging NVM is promising to become the next-generation storage media. However, its high cost hinders its development. Recent deduplication researches in NVM file systems demonstrate that NVM's cost can be reduced by eliminating redundant data blocks, but their design lacks complete insights into NVM's I/O mechanisms.
We propose Light-Dedup, a light-weight inline deduplication framework for NVM file systems that performs fast block-level deduplication while taking NVM's I/O mechanisms into consideration. Specifically, Light-Dedup proposes Light-Redundant-Block-Identifier (LRBI), which combines non-cryptographic hash with a speculative-prefetch-based byte-by-byte content-comparison approach. LRBI leverages the memory interface of NVM to enable asynchronous reads by speculatively prefetching in-NVM data blocks into the CPU/NVM buffers. Thus, NVM's read latency seen by content-comparison is markedly reduced due to buffer hits. Moreover, Light-Dedup adopts an in-NVM Light-Meta-Table (LMT) to store deduplication metadata and collaborate with LRBI. LMT is organized in the region granularity, which significantly reduces metadata I/O amplification and improves deduplication performance.
Experimental results suggest Light-Dedup achieves 1.01-8.98× I/O throughput over the state-of-the-art NVM deduplication file systems. Here, the speculative prefetch technique used in LRBI improves Light-Dedup by 0.3-118%. In addition, the region-based layout of LMT reduces metadata read/write amplification from 19.35×/9.86× to 6.10×/3.43× in our hand-crafted aging workload.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Don't Maintain Twice, It's Alright: Merged Metadata Management in Deduplication File System with GogetaFSYanqi Pan, Wen Xia, Erci Xu, Hao Huang 等FAST 2025 · 被引用 6 次
- XLL: Cross-Layer Logging for Data Deduplication in Consensus-Based StorageJohn Shawger, Arnav Jhingran, Andrea C. Arpaci-Dusseau, Remzi H. Arpaci-DusseauNSDI 2026 · 被引用 1 次
- Fast and Synchronous Crash Consistency with Metadata Write-Once File SystemYanqi Pan, Wen Xia, Yifeng Zhang, Xiangyu Zou 等OSDI 2025
- Towards Condensed and Efficient Read-Only File System via Sort-Enhanced CompressionHao Huang, Yifeng Zhang, Yanqi Pan, Wen Xia 等FAST 2026
它引用的顶会 Paper9
- An Empirical Guide to the Behavior and Use of Scalable Persistent MemoryJian Yang, Juno Kim, Morteza Hoseinzadeh, Joseph Izraelevitz 等FAST 2020 · 被引用 470 次
- Characterizing the performance of intel optane persistent memory: a close look at its on-DIMM bufferingLingfeng Xiang, Xingsheng Zhao, Jia Rao, Song Jiang 等EuroSys 2022 · 被引用 56 次
- Balancing storage efficiency and data confidentiality with tunable encrypted deduplicationJingwei Li, Zuoru Yang, Yanjing Ren, Patrick P. C. Lee 等EuroSys 2020 · 被引用 46 次
- MT^2: Memory Bandwidth Regulation on Hybrid NVM/DRAM PlatformsJifei Yi, Benchao Dong, Mingkai Dong, Ruizhe Tong 等FAST 2022 · 被引用 23 次
- NyxCache: Flexible and Efficient Multi-tenant Persistent Memory CachingKan Wu, Kaiwei Tu, Yuvraj Patel, Rathijit Sen 等FAST 2022 · 被引用 22 次
相关 Paper
- FinerDedup: Sifting Fingerprints for Efficient Data Deduplication on Mobile DevicesXianzhang Chen, Xingjie Zhou, Wei Li, Xi Yu 等DAC 2024 · 被引用 2 次
- ESD: An ECC-assisted and Selective Deduplication for Encrypted Non-Volatile Main MemoryChunfeng Du, Suzhen Wu, Jiapeng Wu, Bo Mao 等HPCA 2023 · 被引用 7 次
- Eliminating Storage Management Overhead of Deduplication over SSD Arrays Through a Hardware/Software Co-DesignYuhong Wen, Xiaogang Zhao, You Zhou, Tong Zhang 等ASPLOS 2024 · 被引用 7 次
- RubbleDB: CPU-Efficient Replication with NVMe-oFHaoyu Li, Sheng Jiang, Chen Chen, Ashwini Raina 等USENIX ATC 2023 · 被引用 9 次
- Austere Flash Caching with Deduplication and CompressionQiuping Wang, Jinhong Li, Wen Xia, Erik Kruus 等USENIX ATC 2020 · 被引用 26 次
