FuseFL: One-Shot Federated Learning through the Lens of Causality with Progressive Model Fusion
Zhenheng Tang, Yonggang Zhang, Peijie Dong, Yiu-ming Cheung, Amelie Chi Zhou, Bo Han, Xiaowen Chu
摘要
One-shot Federated Learning (OFL) significantly reduces communication costs in FL by aggregating trained models only once. However, the performance of advanced OFL methods is far behind the normal FL. In this work, we provide a causal view to find that this performance drop of OFL methods comes from the isolation problem, which means that local isolatedly trained models in OFL may easily fit to spurious correlations due to the data heterogeneity. From the causal perspective, we observe that the spurious fitting can be alleviated by augmenting intermediate features from other clients. Built upon our observation, we propose a novel learning approach to endow OFL with superb performance and low communication and storage costs, termed as FuseFL. Specifically, FuseFL decomposes neural networks into several blocks, and progressively trains and fuses each block following a bottom-up manner for feature augmentation, introducing no additional communication costs. Comprehensive experiments demonstrate that FuseFL outperforms existing OFL and ensemble FL by a significant margin. We conduct comprehensive experiments to show that FuseFL supports high scalability of clients, heterogeneous model training, and low memory costs. Our work is the first attempt using causality to analyze and alleviate data heterogeneity of OFL.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- You Only Communicate Once: One-shot Federated Low-Rank Adaptation of MLLMBinqian Xu, Haiyang Mei, Zechen Bai, Jinjin Gong 等NeurIPS 2025 · 被引用 4 次
- OASIS: One-Shot Federated Graph Learning via Wasserstein Assisted Knowledge IntegrationFrank Wan, Jiaru Qian, Wenke Huang, Qilin Xu 等NeurIPS 2025 · 被引用 2 次
- TOFA: Training-Free One-Shot Federated Adaptation for Vision-Language ModelsLi Zhang, Zhongxuan Han, Xiaohua Feng, Jiaming Zhang 等AAAI 2026 · 被引用 1 次
- Guiding Diffusion Models with Fine-Grained Conditions and Semantics-Preserving Sampling for One-Shot Federated LearningXiaojun Deng, Tianchi Liao, Zhiyuan Liu, Chuan Chen 等CVPR 2026
- Feature-Aware One-Shot Federated Learning via Hierarchical Token SequencesShudong Liu, Hanwen Zhang, Xiuling Wang, Yuesheng Zhu 等AAAI 2026
它引用的顶会 Paper82
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- SCAFFOLD: Stochastic Controlled Averaging for Federated LearningSai Praneeth Karimireddy, Satyen Kale, Mehryar Mohri, Sashank J. Reddi 等ICML 2020 · 被引用 3,875 次
- On the Convergence of FedAvg on Non-IID DataXiang Li, Kaixuan Huang, Wenhao Yang, Shusen Wang 等ICLR 2020 · 被引用 2,930 次
- Tackling the Objective Inconsistency Problem in Heterogeneous Federated OptimizationJianyu Wang, Qinghua Liu, Hao Liang, Gauri Joshi 等NeurIPS 2020 · 被引用 2,231 次
相关 Paper
- Revisiting Ensembling in One-Shot Federated LearningYoussef Allouah, Akash Balasaheb Dhasade, Rachid Guerraoui, Nirupam Gupta 等NeurIPS 2024 · 被引用 21 次
- A Unified Solution to Diverse Heterogeneities in One-Shot Federated LearningJun Bai, Yiliao Song, Di Wu, Atul Sajjanhar 等KDD 2025
- DENSE: Data-Free One-Shot Federated LearningJie Zhang, Chen Chen, Bo Li, Lingjuan Lyu 等NeurIPS 2022 · 被引用 202 次
- One-shot-but-not-degraded Federated LearningHui Zeng, Minrui Xu, Tongqing Zhou, Xinyi Wu 等ACM MM 2024 · 被引用 5 次
- One-shot Federated Learning via Synthetic Distiller-Distillate CommunicationJunyuan Zhang, Songhua Liu, Xinchao WangNeurIPS 2024 · 被引用 21 次
