FedEMoE: Improving Personalization on Heterogeneous Federated Learning via Elastic Mixture of Experts Architecture
Haizhou Du, Lixin Huang, Zonghan Wu, Huan Huo
摘要
Heterogeneous federated learning (HtFL) has emerged as a promising approach to address heterogeneity in local computational resources and data distribution. However, existing methods cause performance degradation of model personalization because personalized and generalized knowledge are either intertwined or dominated by one of them. To address this issue, we propose a novel Elastic Mixture of Experts (EMoE) architecture on HtFL, namely FedEMoE, decoupling personalization from generalization. Specially, FedEMoE employs a multi-scale feature extraction mechanism via personalized experts to enrich personalized knowledge. Furthermore, we design an elastic shared expert to break the transferred knowledge bottleneck across heterogeneous client models. The elastic shared expert can adaptively expand or shrink according to the status of each expert by the weight spectrum analysis, respectively. Extensive experiments across statistical and model heterogeneity settings demonstrate that FedEMoE significantly outperforms state-of-theart methods on the accuracy of each heterogeneous model over diverse datasets.
Building on this insight, our main contributions are:
• To the best of our knowledge, we first design a novel elastic MoE architecture for HtFL to decouple personalization from generalization, namely FedEMoE, ensuring local models never compromise their specialization for global consensus.
• We introduce a multi-scale feature extraction and knowledge exchange mechanism based on a MoE architecture at client-side. This mechanism maximizes the feature captured from local unique data and retained the ability to exchange knowledge with other clients, resulting in the improvement of personalization.
• We design an elastic MoE architecture where the shared expert acts as a dynamic repository of collective intelligence at server-side. Its structure can adaptively strengthen its representation capacity by preserving and integrating specialized knowledge rather than averaging it away.
• We evaluate FedEMoE in settings with model and statistical heterogeneity. Our extensive experiments and ablation studies demonstrate that FedEMoE is superior to state-of-the-art methods, improving task accuracy by up to 48.05%.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper18
- Ensemble Distillation for Robust Model Fusion in Federated LearningTao Lin, Lingjing Kong, Sebastian U. Stich, Martin JaggiNeurIPS 2020 · 被引用 1,615 次
- Federated Learning on Non-IID Data Silos: An Experimental StudyQinbin Li, Yiqun Diao, Quan Chen, Bingsheng HeICDE 2022 · 被引用 1,110 次
- Exploiting Shared Representations for Personalized Federated LearningLiam Collins, Hamed Hassani, Aryan Mokhtari, Sanjay ShakkottaiICML 2021 · 被引用 1,081 次
- Data-Free Knowledge Distillation for Heterogeneous Federated LearningZhuangdi Zhu, Junyuan Hong, Jiayu ZhouICML 2021 · 被引用 957 次
- BASE Layers: Simplifying Training of Large, Sparse ModelsMike Lewis, Shruti Bhosale, Tim Dettmers, Naman Goyal 等ICML 2021 · 被引用 382 次
相关 Paper
- PM-MOE: Mixture of Experts on Private Model Parameters for Personalized Federated LearningYu Feng, Yangli-ao Geng, Yifan Zhu, Zongfu Han 等WWW 2025 · 被引用 12 次
- dFLMoE: Decentralized Federated Learning via Mixture of Experts for Medical Data AnalysisLuyuan Xie, Tianyu Luan, Wenyuan Cai, Guochen Yan 等CVPR 2025
- pFedAFM: Adaptive Feature Mixture for Data-Level Personalization in Heterogeneous Federated Learning on Mobile Edge DevicesLiping Yi, Han Yu, Gang Wang, Xiaoguang Liu 等ICDE 2025 · 被引用 4 次
- pFedES: Generalized Proxy Feature Extractor Sharing for Model Heterogeneous Personalized Federated LearningLiping Yi, Han Yu, Chao Ren, Gang Wang 等AAAI 2025 · 被引用 8 次
- Decoupling General and Personalized Knowledge in Federated Learning via Additive and Low-rank DecompositionXinghao Wu, Xuefeng Liu, Jianwei Niu, Haolin Wang 等ACM MM 2024 · 被引用 15 次
