Exploiting user activeness for data retention in HPC systems
Wei Zhang, Suren Byna, Hyogi Sim, Sangkeun Lee, Sudharshan Vazhkudai, Yong Chen
摘要
HPC systems typically rely on the fixed-lifetime (FLT) data retention strategy, which only considers temporal locality of data accesses to parallel file systems. However, our extensive analysis based on the leadership-class HPC system traces suggests that the FLT approach often fails to capture the dynamics in users' behavior and leads to undesired data purge. In this study, we propose an activeness-based data retention (ActiveDR) solution, which advocates considering the data retention approach from a holistic activeness-based perspective. By evaluating the frequency and impact of users' activities, ActiveDR prioritizes the file purge process for inactive users and rewards active users with extended file lifetime on parallel storage. Our extensive evaluations based on the traces of the prior Titan supercomputer show that, when reaching the same purge target, ActiveDR achieves up to 37% file miss reduction as compared to the current FLT retention methodology.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它相关 Paper
- File System Semantics Requirements of HPC ApplicationsChen Wang, Kathryn Mohror, Marc SnirHPDC 2021 · 被引用 25 次
- Exploit both SMART Attributes and NAND Flash Wear Characteristics to Effectively Forecast SSD-based Storage Failures in ClustersYunfei Gu, Chentao Wu, Xubin HeUSENIX ATC 2024 · 被引用 8 次
- Trident: Task Scheduling over Tiered Storage Systems in Big Data PlatformsHerodotos Herodotou, Elena KakoulliVLDB 2021 · 被引用 10 次
- PACEMAKER: Avoiding HeART attacks in storage clusters with disk-adaptive redundancySaurabh Kadekodi, Francisco Maturana, Suhas Jayaram Subramanya, Juncheng Yang 等OSDI 2020 · 被引用 29 次
- Automating Distributed Tiered Storage Management in Cluster ComputingHerodotos Herodotou, Elena KakoulliVLDB 2020 · 被引用 30 次
