Approximate Gradient Coding for Distributed Learning with Heterogeneous Stragglers
Heekang Song, Wan Choi
摘要
In this paper, we propose an optimally structured gradient coding scheme to mitigate the straggler problem in distributed learning. Conventional gradient coding methods often assume homogeneous straggler models or rely on excessive data replication, limiting performance in real-world heterogeneous systems. To address these limitations, we formulate an optimization problem minimizing residual error while ensuring unbiased gradient estimation by explicitly considering individual straggler probabilities. We derive closed-form solutions for optimal encoding and decoding coefficients via Lagrangian duality and convex optimization, and propose data allocation strategies that reduce both redundancy and computation load. We also analyze convergence behavior for -strongly convex and -smooth loss functions. Numerical results show that our approach significantly reduces the impact of stragglers and accelerates convergence compared to existing methods.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
相关 Paper
- Leveraging partial stragglers within gradient codingAditya Ramamoorthy, Ruoyu Meng, Vrinda S. GirimajiNeurIPS 2024 · 被引用 7 次
- Live Gradient Compensation for Evading Stragglers in Distributed LearningJian Xu, Shao-Lun Huang, Linqi Song, Tian LanINFOCOM 2021 · 被引用 29 次
- Lightweight Projective Derivative Codes for Compressed Asynchronous Gradient DescentPedro Soto, Ilia Ilmer, Haibin Guan, Jun LiICML 2022 · 被引用 3 次
- Sequential Gradient Coding For Straggler MitigationMuralee Nikhil Krishnan, MohammadReza Ebrahimi, Ashish J. KhistiICLR 2023
- Coded Edge ComputingKwang Taik Kim, Carlee Joe-Wong, Mung ChiangINFOCOM 2020 · 被引用 28 次
