Is Aggregation the Only Choice? Federated Learning via Layer-wise Model Recombination
Ming Hu, Zhihao Yue, Xiaofei Xie, Cheng Chen, Yihao Huang, Xian Wei, Xiang Lian, Yang Liu, Mingsong Chen
Abstract
Although Federated Learning (FL) enables global model training Xiaofei Xie xfxie@smu.edu.sg Singapore Management University Singapore, Singapore Xian Wei xwei@sei.ecnu.edu.cn East China Normal University Shanghai, China Mingsong Chen∗ mschen@sei.ecnu.edu.cn East China Normal University Shanghai, China • Computing methodologies → Distributed artificial intelligence. across clients without compromising their raw data, due to the unevenly distributed data among clients, existing Federated Averaging (FedAvg)-based methods suffer from the problem of low inference performance. Specifically, different data distributions among clients lead to various optimization directions of local models. Aggregating local models usually results in a low-generalized global model, which performs worse on most of the clients. To address the above issue, inspired by the observation from a geometric perspective that a well-generalized solution is located in a flat area rather than a sharp area, we propose a novel and heuristic FL paradigm named FedMR (Federated Model Recombination). The goal of FedMR is to guide the recombined models to be trained towards a flat area. Unlike conventional FedAvg-based methods, in FedMR, the cloud server recombines collected local models by shuffling each layer of them to generate multiple recombined models for local training on clients rather than an aggregated global model. Since the area of the f lat area is larger than the sharp area, when local models are located in different areas, recombined models have a higher probability of locating in a flat area. When all recombined models are located in the same flat area, they are optimized towards the same direction. Wetheoretically analyze the convergence of model recombination. Experimental results show that, compared with state-of-the-art FL methods, FedMR can significantly improve the inference accuracy without exposing the privacy of each client.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 607ed5d7-c0a2-4631-9d7f-63b387e2252bCited by top-tier papers19
- Federated Graph Learning under Domain Shift with Generalizable PrototypesGuancheng Wan, Wenke Huang, Mang YeAAAI 2024 · 70 citations
- FIARSE: Model-Heterogeneous Federated Learning via Importance-Aware Submodel ExtractionFeijie Wu, Xingchen Wang, Yaqing Wang, Tianci Liu et al.NeurIPS 2024 · 47 citations
- Parameter Disparities Dissection for Backdoor Defense in Heterogeneous Federated LearningWenke Huang, Mang Ye, Zekun Shi, Guancheng Wan et al.NeurIPS 2024 · 12 citations
- FedGPS: Statistical Rectification Against Data Heterogeneity in Federated LearningZhiqin Yang, Yonggang Zhang, Chenxin Li, Yiu-ming Cheung et al.NeurIPS 2025 · 7 citations
- Gradients as An Action: Towards Communication-Efficient Federated Recommender Systems via Adaptive Action SharingZhufeng Lu, Chentao Jia, Ming Hu, Xiaofei Xie et al.KDD 2025 · 2 citations
Builds on24
- Practical Secure Aggregation for Privacy-Preserving Machine LearningKallista A. Bonawitz, Vladimir Ivanov, Ben Kreuter, Antonio Marcedone et al.CCS 2017 · 3,936 citations
- SCAFFOLD: Stochastic Controlled Averaging for Federated LearningSai Praneeth Karimireddy, Satyen Kale, Mehryar Mohri, Sashank J. Reddi et al.ICML 2020 · 3,875 citations
- On the Convergence of FedAvg on Non-IID DataXiang Li, Kaixuan Huang, Wenhao Yang, Shusen Wang et al.ICLR 2020 · 2,930 citations
- Ensemble Distillation for Robust Model Fusion in Federated LearningTao Lin, Lingjing Kong, Sebastian U. Stich, Martin JaggiNeurIPS 2020 · 1,615 citations
- Optimizing Federated Learning on Non-IID Data with Reinforcement LearningHao Wang, Zakhary Kaplan, Di Niu, Baochun LiINFOCOM 2020 · 1,002 citations
Related papers
- FedCross: Towards Accurate Federated Learning via Multi-Model Cross-AggregationMing Hu, Peiheng Zhou, Zhihao Yue, Zhiwei Ling et al.ICDE 2024 · 32 citations
- FedMut: Generalized Federated Learning via Stochastic MutationMing Hu, Yue Cao, Anran Li, Zhiming Li et al.AAAI 2024 · 46 citations
- FedKDMR: Robust Federated Learning via Joint Knowledge Distillation & Model RecombinationWenhao Li, Christos Anagnostopoulos, Shameem A. Puthiya Parambath, Kevin BrysonKDD 2026
- SemiFL: Semi-Supervised Federated Learning for Unlabeled Clients with Alternate TrainingEnmao Diao, Jie Ding, Vahid TarokhNeurIPS 2022 · 130 citations
- Federated Learning with Manifold Regularization and Normalized Update ReaggregationXuming An, Li Shen, Han Hu, Yong LuoNeurIPS 2023 · 23 citations
