Hybrid Evaluation for Distributed Iterative Matrix Computation
Zihao Chen, Chen Xu, Juan Soto, Volker Markl, Weining Qian, Aoying Zhou
摘要
Distributed matrix computation is common in large-scale data processing and machine learning applications. Existing systems that support distributed matrix computation already explore incremental evaluation for iterative-convergent algorithms. However, they are oblivious to the fact that non-zero increments are scattered in different blocks in a distributed environment. Additionally, we observe that incremental evaluation does not always outperform full evaluation. To address these issues, we propose matrix reorganization to optimize the physical layout upon the state-of-art optimized partition schemes, and thereby accelerate the incremental evaluation. More importantly, we propose a hybrid evaluation to efficiently interleave full and incremental evaluation during the iterative process. In particular, it employs a cost model to compare the overhead costs of two types of evaluations and a selective comparison mechanism to reduce the overhead incurred by comparison itself. To demonstrate the efficiency of our techniques, we implement HyMAC, a hybrid matrix computation system based on SystemML. Our experiments show that HyMAC reduces execution time on large datasets by 23% on average in comparison to the state-of-art optimization technique and consequently outperforms SystemML, ScaLAPACK, and SciDB by an order of magnitude.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- Redundancy Elimination in Distributed Matrix ComputationZihao Chen, Baokun Han, Chen Xu, Weining Qian 等SIGMOD 2022 · 被引用 4 次
- Automatic Optimization of Matrix Implementations for Distributed Machine Learning and Linear AlgebraShangyu Luo, Dimitrije Jankov, Binhang Yuan, Chris JermaineSIGMOD 2021 · 被引用 9 次
- C olumnSGD: A Column-oriented Framework for Distributed Stochastic Gradient DescentZhipeng Zhang, Wentao Wu, Jiawei Jiang, Lele Yu 等ICDE 2020 · 被引用 6 次
- PreVision: An Out-of-Core Matrix Computation System with Optimal Buffer ReplacementKyoseung Koo, Sohyun Kim, Wonhyeon Kim, Yoojin Choi 等SIGMOD 2024 · 被引用 3 次
- A novel data transformation and execution strategy for accelerating sparse matrix multiplication on GPUsPeng Jiang, Changwan Hong, Gagan AgrawalPPoPP 2020 · 被引用 74 次
