FedExP: Speeding Up Federated Averaging via Extrapolation
Divyansh Jhunjhunwala, Shiqiang Wang, Gauri Joshi
摘要
Federated Averaging (FedAvg) remains the most popular algorithm for Federated Learning (FL) optimization due to its simple implementation, stateless nature, and privacy guarantees combined with secure aggregation. Recent work has sought to generalize the vanilla averaging in FedAvg to a generalized gradient descent step by treating client updates as pseudo-gradients and using a server step size. While the use of a server step size has been shown to provide performance improvement theoretically, the practical benefit of the server step size has not been seen in most existing works. In this work, we present FedExP, a method to adaptively determine the server step size in FL based on dynamically varying pseudo-gradients throughout the FL process. We begin by considering the overparameterized convex regime, where we reveal an interesting similarity between FedAvg and the Projection Onto Convex Sets (POCS) algorithm. We then show how FedExP can be motivated as a novel extension to the extrapolation mechanism that is used to speed up POCS. Our theoretical analysis later also discusses the implications of FedExP in underparameterized and non-convex settings. Experimental results show that FedExP consistently converges faster than FedAvg and competing baselines on a range of realistic FL datasets.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper34
- FedCDA: Federated Learning with Cross-rounds Divergence-aware AggregationHaozhao Wang, Haoran Xu, Yichen Li, Yuan Xu 等ICLR 2024 · 被引用 62 次
- Convergence Analysis of Sequential Federated Learning on Heterogeneous DataYipeng Li, Xinchen LyuNeurIPS 2023 · 被引用 53 次
- FIARSE: Model-Heterogeneous Federated Learning via Importance-Aware Submodel ExtractionFeijie Wu, Xingchen Wang, Yaqing Wang, Tianci Liu 等NeurIPS 2024 · 被引用 47 次
- Classifier Clustering and Feature Alignment for Federated Learning under Distributed Concept DriftJunbao Chen, Jingfeng Xue, Yong Wang, Zhenyan Liu 等NeurIPS 2024 · 被引用 31 次
- Federated Learning with Manifold Regularization and Normalized Update ReaggregationXuming An, Li Shen, Han Hu, Yong LuoNeurIPS 2023 · 被引用 23 次
它引用的顶会 Paper11
- SCAFFOLD: Stochastic Controlled Averaging for Federated LearningSai Praneeth Karimireddy, Satyen Kale, Mehryar Mohri, Sashank J. Reddi 等ICML 2020 · 被引用 3,875 次
- Tackling the Objective Inconsistency Problem in Heterogeneous Federated OptimizationJianyu Wang, Qinghua Liu, Hao Liang, Gauri Joshi 等NeurIPS 2020 · 被引用 2,231 次
- Adaptive Federated OptimizationSashank J. Reddi, Zachary Charles, Manzil Zaheer, Zachary Garrett 等ICLR 2021 · 被引用 1,917 次
- ProxSkip: Yes! Local Gradient Steps Provably Lead to Communication Acceleration! Finally!Konstantin Mishchenko, Grigory Malinovsky, Sebastian U. Stich, Peter RichtárikICML 2022 · 被引用 200 次
- Linear Convergence in Federated Learning: Tackling Client Heterogeneity and Sparse GradientsAritra Mitra, Rayana H. Jaafar, George J. Pappas, Hamed HassaniNeurIPS 2021 · 被引用 193 次
相关 Paper
- The Power of Extrapolation in Federated LearningHanmin Li, Kirill Acharya, Peter RichtárikNeurIPS 2024 · 被引用 16 次
- P-FedAvg: Parallelizing Federated Learning with Theoretical GuaranteesZhicong Zhong, Yipeng Zhou, Di Wu, Xu Chen 等INFOCOM 2021 · 被引用 67 次
- A New Theoretical Perspective on Data Heterogeneity in Federated OptimizationJiayi Wang, Shiqiang Wang, Rong-Rong Chen, Mingyue JiICML 2024 · 被引用 3 次
- A Lightweight Method for Tackling Unknown Participation Statistics in Federated AveragingShiqiang Wang, Mingyue JiICLR 2024
- Personalized Federated Learning with Moreau EnvelopesCanh T. Dinh, Nguyen Hoang Tran, Tuan Dung NguyenNeurIPS 2020 · 被引用 1,542 次
