FedNew: A Communication-Efficient and Privacy-Preserving Newton-Type Method for Federated Learning
Anis Elgabli, Chaouki Ben Issaid, Amrit Singh Bedi, Ketan Rajawat, Mehdi Bennis, Vaneet Aggarwal
Abstract
Newton-type methods are popular in federated learning due to their fast convergence. Still, they suffer from two main issues, namely: low communication efficiency and low privacy due to the requirement of sending Hessian information from clients to parameter server (PS). In this work, we introduced a novel framework called FedNew in which there is no need to transmit Hessian information from clients to PS, hence resolving the bottleneck to improve communication efficiency. In addition, FedNew hides the gradient information and results in a privacy-preserving approach compared to the existing state-of-the-art. The core novel idea in FedNew is to introduce a two level framework, and alternate between updating the inverse Hessian-gradient product using only one alternating direction method of multipliers (ADMM) step and then performing the global model update using Newton's method. Though only one ADMM pass is used to approximate the inverse Hessian-gradient product at each iteration, we develop a novel theoretical approach to show the converging behavior of FedNew for convex problems. Additionally, a significant reduction in communication overhead is achieved by utilizing stochastic quantization. Numerical results using real datasets show the superiority of FedNew compared to existing methods in terms of communication costs.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers8
- Out-of-Distribution Generalization of Federated Learning via Implicit Invariant RelationshipsYaming Guo, Kai Guo, Xiaofeng Cao, Tieru Wu et al.ICML 2023 · 46 citations
- Improved Communication Efficiency in Federated Natural Policy Gradient via ADMM-based Gradient UpdatesGuangchen Lan, Han Wang, James Anderson, Christopher G. Brinton et al.NeurIPS 2023 · 32 citations
- Federated Combinatorial Multi-Agent Multi-Armed BanditsFares Fourati, Mohamed-Slim Alouini, Vaneet AggarwalICML 2024 · 10 citations
- FedMuon: Federated Learning with Bias-corrected LMO-based OptimizationYuki Takezawa, Anastasia Koloskova, Xiaowen Jiang, Sebastian U. StichICLR 2026 · 9 citations
- FedNS: A Fast Sketching Newton-Type Algorithm for Federated LearningJian Li, Yong Liu, Weiping WangAAAI 2024 · 7 citations
Builds on6
- Adaptive Federated OptimizationSashank J. Reddi, Zachary Charles, Manzil Zaheer, Zachary Garrett et al.ICLR 2021 · 1,917 citations
- Deep Models Under the GAN: Information Leakage from Collaborative Deep LearningBriland Hitaj, Giuseppe Ateniese, Fernando Pérez-CruzCCS 2017 · 1,581 citations
- Adaptive Gradient Descent without DescentYura Malitsky, Konstantin MishchenkoICML 2020 · 171 citations
- Acceleration for Compressed Gradient Descent in Distributed and Federated OptimizationZhize Li, Dmitry Kovalev, Xun Qian, Peter RichtárikICML 2020 · 156 citations
- FedNL: Making Newton-Type Methods Applicable to Federated LearningMher Safaryan, Rustem Islamov, Xun Qian, Peter RichtárikICML 2022 · 90 citations
Related papers
- Resolving the Tug-of-War: A Separation of Communication and Learning in Federated LearningJunyi Li, Heng HuangNeurIPS 2023 · 3 citations
- Communication-Efficient Adaptive Federated LearningYujia Wang, Lu Lin, Jinghui ChenICML 2022 · 101 citations
- Masked Random Noise for Communication-Efficient Federated LearningShiwei Li, Yingyi Cheng, Haozhao Wang, Xing Tang et al.ACM MM 2024 · 7 citations
- Newton-ADMM: a distributed GPU-accelerated optimizer for multiclass classification problemsChih-Hao Fang, Sudhir B. Kylasa, Fred Roosta, Michael W. Mahoney et al.SC 2020 · 5 citations
- DAdaQuant: Doubly-adaptive quantization for communication-efficient Federated LearningRobert Hönig, Yiren Zhao, Robert MullinsICML 2022 · 87 citations
