Fast Federated Learning in the Presence of Arbitrary Device Unavailability
Xinran Gu, Kaixuan Huang, Jingzhao Zhang, Longbo Huang
Abstract
Federated Learning (FL) coordinates with numerous heterogeneous devices to collaboratively train a shared model while preserving user privacy. Despite its multiple advantages, FL faces new challenges. One challenge arises when devices drop out of the training process beyond the control of the central server. In this case, the convergence of popular FL algorithms such as FedAvg is severely influenced by the straggling devices. To tackle this challenge, we study federated learning algorithms under arbitrary device unavailability and propose an algorithm named Memory-augmented Impatient Federated Averaging (MIFA). Our algorithm efficiently avoids excessive latency induced by inactive devices, and corrects the gradient bias using the memorized latest updates from the devices. We prove that MIFA achieves minimax optimal convergence rates on non-i.i.d. data for both strongly convex and non-convex smooth functions. We also provide an explicit characterization of the improvement over baseline algorithms through a case study, and validate the results by numerical experiments on real-world datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6cbe3721-8d89-47ee-b2e7-da127c8a2053Cited by top-tier papers25
- Sharper Convergence Guarantees for Asynchronous SGD for Distributed and Federated LearningAnastasia Koloskova, Sebastian U. Stich, Martin JaggiNeurIPS 2022 · 131 citations
- Communication-Efficient Device Scheduling for Federated Learning Using Stochastic OptimizationJake B. Perazzone, Shiqiang Wang, Mingyue Ji, Kevin S. ChanINFOCOM 2022 · 88 citations
- A Unified Analysis of Federated Learning with Arbitrary Client ParticipationShiqiang Wang, Mingyue JiNeurIPS 2022 · 85 citations
- Federated Minimax Optimization: Improved Convergence Analyses and AlgorithmsPranay Sharma, Rohan Panda, Gauri Joshi, Pramod K. VarshneyICML 2022 · 63 citations
- Anarchic Federated LearningHaibo Yang, Xin Zhang, Prashant Khanduri, Jia LiuICML 2022 · 62 citations
Builds on5
- SCAFFOLD: Stochastic Controlled Averaging for Federated LearningSai Praneeth Karimireddy, Satyen Kale, Mehryar Mohri, Sashank J. Reddi et al.ICML 2020 · 3,875 citations
- On the Convergence of FedAvg on Non-IID DataXiang Li, Kaixuan Huang, Wenhao Yang, Shusen Wang et al.ICLR 2020 · 2,930 citations
- Tackling the Objective Inconsistency Problem in Heterogeneous Federated OptimizationJianyu Wang, Qinghua Liu, Hao Liang, Gauri Joshi et al.NeurIPS 2020 · 2,231 citations
- Why are Adaptive Methods Good for Attention Models?Jingzhao Zhang, Sai Praneeth Karimireddy, Andreas Veit, Seungyeon Kim et al.NeurIPS 2020 · 397 citations
- Achieving Linear Speedup with Partial Worker Participation in Non-IID Federated LearningHaibo Yang, Minghong Fang, Jia LiuICLR 2021 · 310 citations
Related papers
- Efficient Federated Learning against Heterogeneous and Non-stationary Client UnavailabilityMing Xiang, Stratis Ioannidis, Edmund Yeh, Carlee Joe-Wong et al.NeurIPS 2024 · 26 citations
- Federated Learning under Heterogeneous and Correlated Client AvailabilityAngelo Rodio, Francescomaria Faticanti, Othmane Marfoq, Giovanni Neglia et al.INFOCOM 2023 · 27 citations
- Debiasing Federated Learning with Correlated Client ParticipationZhenyu Sun, Ziyang Zhang, Zheng Xu, Gauri Joshi et al.ICLR 2025
- Delayed Gradient Averaging: Tolerate the Communication Latency for Federated LearningLigeng Zhu, Hongzhou Lin, Yao Lu, Yujun Lin et al.NeurIPS 2021 · 4 citations
- Towards Straggler-Resilient Split Federated Learning: An Unbalanced Update ApproachDandan Liang, Jianing Zhang, Evan Chen, Zhe Li et al.NeurIPS 2025 · 8 citations
