On the Convergence of Communication-Efficient Local SGD for Federated Learning
Hongchang Gao, An Xu, Heng Huang
Abstract
Federated Learning (FL) has attracted increasing attention in recent years. A leading training algorithm in FL is local SGD, which updates the model parameter on each worker and averages model parameters across different workers only once in a while. Although it has fewer communication rounds than the classical parallel SGD, local SGD still has large communication overhead in each communication round for large machine learning models, such as deep neural networks. To address this issue, we propose a new communicationefficient distributed SGD method, which can significantly reduce the communication cost by the error-compensated double compression mechanism. Under the non-convex setting, our theoretical results show that our approach has better communication complexity than existing methods and enjoys the same linear speedup regarding the number of workers as the full-precision local SGD. Moreover, we propose a communication-efficient distributed SGD with momentum, which also has better communication complexity than existing methods and enjoys a linear speedup with respect to the number of workers. At last, extensive experiments are conducted to verify the performance of our proposed methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 759efbe6-7641-452b-89ae-0462fff569ebCited by top-tier papers14
- Stochastic Controlled Averaging for Federated Learning with Communication CompressionXinmeng Huang, Ping Li, Xiaoyun LiICLR 2024 · 288 citations
- Accelerated Federated Learning with Decoupled Adaptive OptimizationJiayin Jin, Jiaxiang Ren, Yang Zhou, Lingjuan Lyu et al.ICML 2022 · 62 citations
- Optimal Rate Adaption in Federated Learning with Compressed CommunicationsLaizhong Cui, Xiaoxin Su, Yipeng Zhou, Jiangchuan LiuINFOCOM 2022 · 61 citations
- Closing the Generalization Gap of Cross-silo Federated Medical Image SegmentationAn Xu, Wenqi Li, Pengfei Guo, Dong Yang et al.CVPR 2022 · 58 citations
- Coordinating Momenta for Cross-Silo Federated LearningAn Xu, Heng HuangAAAI 2022 · 24 citations
Related papers
- Communication-Efficient Adaptive Federated LearningYujia Wang, Lu Lin, Jinghui ChenICML 2022 · 101 citations
- Step-Ahead Error Feedback for Distributed Training with Compressed GradientAn Xu, Zhouyuan Huo, Heng HuangAAAI 2021 · 17 citations
- Detached Error Feedback for Distributed SGD with Random SparsificationAn Xu, Heng HuangICML 2022 · 12 citations
- Hybrid Local SGD for Federated Learning with Heterogeneous CommunicationsYuanxiong Guo, Ying Sun, Rui Hu, Yanmin GongICLR 2022 · 63 citations
- FedMoS: Taming Client Drift in Federated Learning with Double Momentum and Adaptive SelectionXiong Wang, Yuxin Chen, Yuqing Li, Xiaofei Liao et al.INFOCOM 2023 · 13 citations
