Delayed Gradient Averaging: Tolerate the Communication Latency for Federated Learning
Ligeng Zhu, Hongzhou Lin, Yao Lu, Yujun Lin, Song Han
摘要
Federated Learning is an emerging direction in distributed machine learning that enables jointly training a model without sharing the data. Since the data is distributed across many edge devices through wireless / long-distance connections, federated learning suffers from inevitable high communication latency. However, the latency issues are undermined in the current literature [15] and existing approaches such as FedAvg [27] become less efficient when the latency increases. To overcome the problem, we propose Delayed Gradient Averaging (DGA), which delays the averaging step to improve efficiency and allows local computation in parallel to communication. We theoretically prove that DGA attains a similar convergence rate as FedAvg, and empirically show that our algorithm can tolerate high network latency without compromising accuracy. Specifically, we benchmark the training speed on various vision (CIFAR, ImageNet) and language tasks (Shakespeare), with both IID and non-IID partitions, and show DGA can bring 2.55⇥ to 4.07⇥ speedup. Moreover, we built a 16-node Raspberry Pi cluster and show that DGA can consistently speed up real-world federated learning applications.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- Spectral Co-Distillation for Personalized Federated LearningZihan Chen, Howard H. Yang, Tony Q. S. Quek, Kai Fong Ernest ChongNeurIPS 2023 · 被引用 30 次
- MAS: Towards Resource-Efficient Federated Multiple-Task LearningWeiming Zhuang, Yonggang Wen, Lingjuan Lyu, Shuai ZhangICCV 2023 · 被引用 22 次
- No One Idles: Efficient Heterogeneous Federated Learning with Parallel Edge and Server ComputationFeilong Zhang, Xianming Liu, Shiyi Lin, Gang Wu 等ICML 2023 · 被引用 15 次
- Fedhca2: Towards Hetero-Client Federated Multi-Task LearningYuxiang Lu, Suizhi Huang, Yuwen Yang, Shalayiding Sirejiding 等CVPR 2024 · 被引用 13 次
- Workie-Talkie: Accelerating Federated Learning by Overlapping Computing and Communications via Contrastive RegularizationRui Chen, Qiyu Wan, Pavana Prakash, Lan Zhang 等ICCV 2023 · 被引用 10 次
它引用的顶会 Paper6
- SCAFFOLD: Stochastic Controlled Averaging for Federated LearningSai Praneeth Karimireddy, Satyen Kale, Mehryar Mohri, Sashank J. Reddi 等ICML 2020 · 被引用 3,875 次
- On the Convergence of FedAvg on Non-IID DataXiang Li, Kaixuan Huang, Wenhao Yang, Shusen Wang 等ICLR 2020 · 被引用 2,930 次
- Adaptive Federated OptimizationSashank J. Reddi, Zachary Charles, Manzil Zaheer, Zachary Garrett 等ICLR 2021 · 被引用 1,917 次
- Fair Resource Allocation in Federated LearningTian Li, Maziar Sanjabi, Ahmad Beirami, Virginia SmithICLR 2020 · 被引用 971 次
- FetchSGD: Communication-Efficient Federated Learning with SketchingDaniel Rothchild, Ashwinee Panda, Enayat Ullah, Nikita Ivkin 等ICML 2020 · 被引用 425 次
相关 Paper
- AOCC-FL: Federated Learning with Aligned Overlapping via Calibrated CompensationHaozhao Wang, Wenchao Xu, Yunfeng Fan, Ruixuan Li 等INFOCOM 2023 · 被引用 8 次
- FedADMM: A Robust Federated Deep Learning Framework with Adaptivity to System HeterogeneityYonghai Gong, Yichuan Li, Nikolaos M. FrerisICDE 2022 · 被引用 41 次
- Hybrid Local SGD for Federated Learning with Heterogeneous CommunicationsYuanxiong Guo, Ying Sun, Rui Hu, Yanmin GongICLR 2022 · 被引用 63 次
- Enhancing Decentralized Federated Learning for Non-IID Data on Heterogeneous DevicesMin Chen, Yang Xu, Hongli Xu, Liusheng HuangICDE 2023 · 被引用 25 次
- FSL-SAGE: Accelerating Federated Split Learning via Smashed Activation Gradient EstimationSrijith Nair, Michael Lin, Peizhong Ju, Amirreza Talebi 等ICML 2025
