Blaze: Holistic Caching for Iterative Data Processing
Won Wook Song, Jeongyoon Eo, Taegeon Um, Myeongjae Jeon, Byung-Gon Chun
摘要
Modern data processing workloads, such as machine learning and graph processing, involve iterative computations to converge generated models into higher accuracy. An effective caching mechanism is vital to expedite iterative computations since the intermediate data that needs to be stored in memory grows larger over iterations, often exceeding the memory capacity. However, existing systems handle intermediate data through separate operational layers (e.g., caching, eviction, and recovery), with each layer working independently in a greedy or cost-agnostic manner. These layers typically rely on user annotations and past access patterns, failing to make globally optimal decisions for the workload.
To overcome these limitations, Blaze introduces a unified caching mechanism that integrates the separate operational layers. Blaze dynamically captures the workload structure and metrics using profiling and inductive regression, and automatically estimates the potential data caching efficiency associated with different operational decisions based on the profiled information. To achieve this goal, Blaze incorporates potential data recovery costs across stages into a single cost optimization function, which informs the optimal partition state and location. This approach reduces the significant disk I/O overheads caused by oversized partitions and the recomputation overheads for partitions with long lineages, while efficiently utilizing the constrained memory space. Our evaluations demonstrate that Blaze can accelerate end-to-end application completion time by 2.02 -2.52× compared to recomputation-based MEM_ONLY Spark, and by 1.08 -2.86× compared to checkpoint-based MEM+DISK Spark, while reducing the cache data stored on disk by 95% on average.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper7
- Learning Relaxed Belady for Content Distribution Network CachingZhenyu Song, Daniel S. Berger, Kai Li, Wyatt LloydNSDI 2020 · 被引用 193 次
- Capuchin: Tensor-based GPU Memory Management for Deep LearningXuan Peng, Xuanhua Shi, Hulin Dai, Hai Jin 等ASPLOS 2020 · 被引用 143 次
- Caching with Delayed HitsNirav Atre, Justine Sherry, Weina Wang, Daniel S. BergerSIGCOMM 2020 · 被引用 47 次
- Echo: Compiler-based GPU Memory Footprint Reduction for LSTM RNN TrainingBojian Zheng, Nandita Vijaykumar, Gennady PekhimenkoISCA 2020 · 被引用 34 次
- Sponge: Fast Reactive Scaling for Stream Processing with Serverless FrameworksWon Wook Song, Taegeon Um, Sameh Elnikety, Myeongjae Jeon 等USENIX ATC 2023 · 被引用 28 次
相关 Paper
- Blaze: Fast Graph Processing on Fast SSDsJuno Kim, Steven SwansonSC 2022 · 被引用 8 次
- Juggler: Autonomous Cost Optimization and Performance Prediction of Big Data ApplicationsHani Al-Sayeh, Bunjamin Memishi, Muhammad Attahir Jibril, Marcus Paradies 等SIGMOD 2022 · 被引用 16 次
- A Community Cache with Complete InformationMania Abdi, Amin Mosayyebzadeh, Mohammad Hossein Hajkazemi, Emine Ugur Kaynar 等FAST 2021 · 被引用 2 次
- GoCache: Accelerating Out-Of-Core Graph Queries with Pattern-Driven CachingZheng Yang, Yicheng Zhang, Lixiao Cui, Luofan Chen 等ICDE 2026
- uCache: A Customizable Unikernel-based IO CacheIlya Meignan-Masson, Masanori Misono, Viktor Leis, Pramod BhatotiaFAST 2026 · 被引用 1 次
