A Flexible Framework for Communication-Efficient Machine Learning
Sarit Khirirat, Sindri Magnússon, Arda Aytekin, Mikael Johansson
摘要
With the increasing scale of machine learning tasks, it has become essential to reduce the communication between computing nodes. Early work on gradient compression focused on the bottleneck between CPUs and GPUs, but communication-efficiency is now needed in a variety of different system architectures, from high-performance clusters to energy-constrained IoT devices. In the current practice, compression levels are typically chosen before training and settings that work well for one task may be vastly suboptimal for another dataset on another architecture. In this paper, we propose a flexible framework which adapts the compression level to the true gradient at each iteration, maximizing the improvement in the objective function that is achieved per communicated bit. Our framework is easy to adapt from one technology to the next by modeling how the communication cost depends on the compression level for the specific technology. Theoretical results and practical experiments indicate that the automatic tuning strategies significantly increase communication efficiency on several state-of-the-art compression schemes.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Distributed Extra-gradient with Optimal Complexity and Communication GuaranteesAli Ramezani-Kebrya, Kimon Antonakopoulos, Igor Krawczuk, Justin Deschenaux 等ICLR 2023
- LEGACY: A Lightweight Dynamic Gradient Compression Strategy for Distributed Deep LearningMostapha Essoullami, El Houcine Bergou, Aritra DuttaICLR 2026
它引用的顶会 Paper2
相关 Paper
- Tail: An Automated and Lightweight Gradient Compression Framework for Distributed Deep LearningJinrong Guo, Songlin Hu, Wang Wang, Chunrong Yao 等DAC 2020 · 被引用 3 次
- DAGC: Data-Aware Adaptive Gradient CompressionRongwei Lu, Jiajun Song, Bin Chen, Laizhong Cui 等INFOCOM 2023 · 被引用 12 次
- On the Discrepancy between the Theoretical Analysis and Practical Implementations of Compressed Communication for Distributed Deep LearningAritra Dutta, El Houcine Bergou, Ahmed M. Abdelmoniem, Chen-Yu Ho 等AAAI 2020
- DC2: Delay-aware Compression Control for Distributed Machine LearningAhmed M. Abdelmoniem, Marco CaniniINFOCOM 2021 · 被引用 33 次
- Adaptive Gradient Quantization for Data-Parallel SGDFartash Faghri, Iman Tabrizian, Ilia Markov, Dan Alistarh 等NeurIPS 2020 · 被引用 108 次
