Layer-Wise Adaptive Model Aggregation for Scalable Federated Learning
Sunwoo Lee, Tuo Zhang, Amir Salman Avestimehr
Abstract
In Federated Learning, a common approach for aggregating local models across clients is periodic averaging of the full model parameters. It is, however, known that different layers of neural networks can have a different degree of model discrepancy across the clients. The conventional full aggregation scheme does not consider such a difference and synchronizes the whole model parameters at once, resulting in inefficient network bandwidth consumption. Aggregating the parameters that are similar across the clients does not make meaningful training progress while increasing the communication cost. We propose FedLAMA, a layer-wise model aggregation scheme for scalable Federated Learning. FedLAMA adaptively adjusts the aggregation interval in a layer-wise manner, jointly considering the model discrepancy and the communication cost. The layer-wise aggregation method enables to finely control the aggregation interval to relax the aggregation frequency without a significant impact on the model accuracy. Our extensive empirical study shows that, as the model aggregation interval increases, FedLAMA shows a significantly smaller accuracy drop than the periodic full aggregation scheme while achieving comparable communication efficiency.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 907e3417-6218-45d9-8e36-d7cf06898cecCited by top-tier papers12
- LoRA-FAIR: Federated LoRA Fine-Tuning with Aggregation and Initialization RefinementJieming Bian, Lei Wang, Letian Zhang, Jie XuICCV 2025 · 56 citations
- FedLF: Layer-Wise Fair Federated LearningZibin Pan, Chi Li, Fangchen Yu, Shuyi Wang et al.AAAI 2024 · 12 citations
- Layer-wise Update Aggregation with Recycling for Communication-Efficient Federated LearningJisoo Kim, Sungmin Kang, Sunwoo LeeNeurIPS 2025 · 4 citations
- FedDifRC: Unlocking the Potential of Text-to-Image Diffusion Models in Heterogeneous Federated LearningHuan Wang, Haoran Li, Huaming Chen, Jun Yan et al.ICCV 2025 · 3 citations
- Layer-Wise Adaptive Gradient Norm Penalizing Method for Efficient and Accurate Deep LearningSunwoo LeeKDD 2024 · 2 citations
Builds on6
- SCAFFOLD: Stochastic Controlled Averaging for Federated LearningSai Praneeth Karimireddy, Satyen Kale, Mehryar Mohri, Sashank J. Reddi et al.ICML 2020 · 3,875 citations
- Tackling the Objective Inconsistency Problem in Heterogeneous Federated OptimizationJianyu Wang, Qinghua Liu, Hao Liang, Gauri Joshi et al.NeurIPS 2020 · 2,231 citations
- Large Batch Optimization for Deep Learning: Training BERT in 76 minutesYang You, Jing Li, Sashank J. Reddi, Jonathan Hseu et al.ICLR 2020 · 1,170 citations
- HeteroFL: Computation and Communication Efficient Federated Learning for Heterogeneous ClientsEnmao Diao, Jie Ding, Vahid TarokhICLR 2021 · 179 citations
- Accelerating Training of Transformer-Based Language Models with Progressive Layer DroppingMinjia Zhang, Yuxiong HeNeurIPS 2020 · 126 citations
Related papers
- Why Go Full? Elevating Federated Learning Through Partial Network UpdatesHaolin Wang, Xuefeng Liu, Jianwei Niu, Wenkai Guo et al.NeurIPS 2024 · 12 citations
- FedPara: Low-rank Hadamard Product for Communication-Efficient Federated LearningNam Hyeon-Woo, Moon Ye-Bin, Tae-Hyun OhICLR 2022 · 179 citations
- FedLWS: Federated Learning with Adaptive Layer-wise Weight ShrinkingChanglong Shi, Jinmeng Li, He Zhao, Dandan Guo et al.ICLR 2025
- Layer-wised Model Aggregation for Personalized Federated LearningXiaosong Ma, Jie Zhang, Song Guo, Wenchao XuCVPR 2022 · 212 citations
- Elastic Aggregation for Federated OptimizationDengsheng Chen, Jie Hu, Vince Junkai Tan, Xiaoming Wei et al.CVPR 2023
