PM-MOE: Mixture of Experts on Private Model Parameters for Personalized Federated Learning
Yu Feng, Yangli-ao Geng, Yifan Zhu, Zongfu Han, Xie Yu, Kaiwen Xue, Haoran Luo, Mengyang Sun, Guangwei Zhang, Meina Song
Abstract
Federated learning (FL) has gained widespread attention for its privacy-preserving and collaborative learning capabilities. Due to significant statistical heterogeneity, traditional FL struggles to generalize a shared model across diverse data domains. Personalized federated learning addresses this issue by dividing the model into a globally shared part and a locally private part, with the local model correcting representation biases introduced by the global model. Nevertheless, locally converged parameters more accurately capture domain-specific knowledge, and current methods overlook the potential benefits of these parameters. To address these limitations, we propose PM-MoE architecture. This architecture integrates a mixture of personalized modules and an energy-based personalized modules denoising, enabling each client to select beneficial personalized parameters from other clients. We applied the PM-MoE architecture to nine recent model-split-based personalized federated learning algorithms, achieving performance improvements with minimal additional training. Extensive experiments on six widely adopted datasets and two heterogeneity settings validate the effectiveness of our approach. The source code is available at https://github.com/dannis97500/PM-MOE.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers4
- Personalized Additive Modeling for Multi-level Federated LearningShutong Chen, Guodong Long, Tianyi Zhou, Jie Ma et al.ICML 2026 · 2 citations
- FedMerge: Federated Model Merging for PersonalizationShutong Chen, Tianyi Zhou, Guodong Long, Jing Jiang et al.AAAI 2026 · 2 citations
- Aligning by Misaligning: Boundary-aware Curriculum Learning for Multimodal AlignmentHua Ye, Hang Ding, Siyuan Chen, Yiyang Jiang et al.NeurIPS 2025
- Double-Filter: Efficient Fine-tuning of Pre-trained Vision-Language Models via Patch&Layer FilteringYaoqin He, Junchen Fu, Kaiwen Zheng, Songpei Xu et al.ICML 2025
Builds on24
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida et al.NeurIPS 2022 · 24,707 citations
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao et al.ICCV 2023 · 13,211 citations
- SCAFFOLD: Stochastic Controlled Averaging for Federated LearningSai Praneeth Karimireddy, Satyen Kale, Mehryar Mohri, Sashank J. Reddi et al.ICML 2020 · 3,875 citations
- Energy-based Out-of-distribution DetectionWeitang Liu, Xiaoyun Wang, John D. Owens, Yixuan LiNeurIPS 2020 · 2,213 citations
- Ensemble Distillation for Robust Model Fusion in Federated LearningTao Lin, Lingjing Kong, Sebastian U. Stich, Martin JaggiNeurIPS 2020 · 1,615 citations
Related papers
- FedEMoE: Improving Personalization on Heterogeneous Federated Learning via Elastic Mixture of Experts ArchitectureHaizhou Du, Lixin Huang, Zonghan Wu, Huan HuoICML 2026
- Decoupling General and Personalized Knowledge in Federated Learning via Additive and Low-rank DecompositionXinghao Wu, Xuefeng Liu, Jianwei Niu, Haolin Wang et al.ACM MM 2024 · 15 citations
- pFedES: Generalized Proxy Feature Extractor Sharing for Model Heterogeneous Personalized Federated LearningLiping Yi, Han Yu, Chao Ren, Gang Wang et al.AAAI 2025 · 8 citations
- GPFL: Simultaneously Learning Global and Personalized Feature Information for Personalized Federated LearningJianqing Zhang, Yang Hua, Hao Wang, Tao Song et al.ICCV 2023 · 73 citations
- Eliminating Domain Bias for Federated Learning in Representation SpaceJianqing Zhang, Yang Hua, Jian Cao, Hao Wang et al.NeurIPS 2023 · 105 citations
