Low Precision Local Training is Enough for Federated Learning
Zhiwei Li, Yiqiu LI, Binbin Lin, Zhongming Jin, Weizhong Zhang
摘要
Federated Learning (FL) is a prevalent machine learning paradigm designed to address challenges posed by heterogeneous client data while preserving data privacy. Unlike distributed training, it typically orchestrates resource-constrained edge devices to communicate via a low-bandwidth communication network with a central server. This urges the development of more computation and communication efficient training algorithms. In this paper, we propose an efficient FL paradigm, where the local models in the clients are trained with low-precision operations and communicated with the server in low precision format, while only the model aggregation in the server is performed with high-precision computation. We surprisingly find that high precision models can be recovered from the low precision local models with proper aggregation in the server. In this way, both the workload in the client-side and the communication cost can be significantly reduced. We theoretically show that our proposed paradigm can converge to the optimal solution as the training goes on, which demonstrates that low precision local training is enough for FL. Our paradigm can be integrated with existing FL algorithms flexibly. Experiments across extensive benchmarks are conducted to showcase the effectiveness of our proposed method. Notably, the models trained by our method with the precision as low as 8 bits are comparable to those from the full precision training. As a by-product, we show that low precision local training can relieve the over-fitting issue in local training, which under heterogeneous client data can cause the client models drift further away from each other and lead to the failure in model aggregation. Code is released at https://github.com/digbangbang/LPT-FL .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper15
- SCAFFOLD: Stochastic Controlled Averaging for Federated LearningSai Praneeth Karimireddy, Satyen Kale, Mehryar Mohri, Sashank J. Reddi 等ICML 2020 · 被引用 3,875 次
- On the Convergence of FedAvg on Non-IID DataXiang Li, Kaixuan Huang, Wenhao Yang, Shusen Wang 等ICLR 2020 · 被引用 2,930 次
- Tackling the Objective Inconsistency Problem in Heterogeneous Federated OptimizationJianyu Wang, Qinghua Liu, Hao Liang, Gauri Joshi 等NeurIPS 2020 · 被引用 2,231 次
- Ensemble Distillation for Robust Model Fusion in Federated LearningTao Lin, Lingjing Kong, Sebastian U. Stich, Martin JaggiNeurIPS 2020 · 被引用 1,615 次
- Exploiting Shared Representations for Personalized Federated LearningLiam Collins, Hamed Hassani, Aryan Mokhtari, Sanjay ShakkottaiICML 2021 · 被引用 1,081 次
相关 Paper
- FedLMT: Tackling System Heterogeneity of Federated Learning via Low-Rank Model Training with Theoretical GuaranteesJiahao Liu, Yipeng Zhou, Di Wu, Miao Hu 等ICML 2024 · 被引用 8 次
- Resource-Efficient Federated Learning with Hierarchical Aggregation in Edge ComputingZhiyuan Wang, Hongli Xu, Jianchun Liu, He Huang 等INFOCOM 2021 · 被引用 216 次
- No One Idles: Efficient Heterogeneous Federated Learning with Parallel Edge and Server ComputationFeilong Zhang, Xianming Liu, Shiyi Lin, Gang Wu 等ICML 2023 · 被引用 15 次
- Towards Energy-efficient Federated Learning via INT8-based Training on Mobile DSPsJinliang Yuan, Shangguang Wang, Hongyu Li, Daliang Xu 等WWW 2024 · 被引用 8 次
- Resolving the Tug-of-War: A Separation of Communication and Learning in Federated LearningJunyi Li, Heng HuangNeurIPS 2023 · 被引用 3 次
