FedAFD: Multimodal Federated Learning via Adversarial Fusion and Distillation
Min Tan, Junchao Ma, Yinfu FENG, Jiajun Ding, Wenwen Pan, Tingting Han, Qian Zheng, Zhenzhong Kuang, Zhou Yu
Abstract
Multimodal Federated Learning (MFL) enables clients with heterogeneous data modalities to collaboratively train models without sharing raw data, offering a privacy-preserving framework that leverages complementary cross-modal information. However, existing methods often overlook personalized client performance and struggle with modality/task discrepancies, as well as model heterogeneity. To address these challenges, we propose FedAFD, a unified MFL framework that enhances client and server learning. On the client side, we introduce a bi-level adversarial alignment strategy to align local and global representations within and across modalities, mitigating modality and task gaps. We further design a granularity-aware fusion module to integrate global knowledge into the personalized features adaptively. On the server side, to handle model heterogeneity, we propose a similarity-guided ensemble distillation mechanism that aggregates client representations on shared public data based on feature similarity and distills the fused knowledge into the global model. Extensive experiments conducted under both IID and non-IID settings demonstrate that FedAFD 1 achieves superior performance and efficiency for both the client and the server.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3ceeab97-3c01-433f-bf20-63314d83f9a0Builds on22
- Hybrid CNN-Transformer Feature Fusion for Single Image DerainingXiang Chen, Jinshan Pan, Jiyang Lu, Zhentao Fan et al.AAAI 2023 · 75 citations
- Feature Fusion from Head to Tail for Long-Tailed Visual RecognitionMengke Li, Zhikai Hu, Yang Lu, Weichao Lan et al.AAAI 2024 · 59 citations
- Enriching Multimodal Sentiment Analysis Through Textual Emotional Descriptions of Visual-Audio ContentSheng Wu, Dongxiao He, Xiaobao Wang, Longbiao Wang et al.AAAI 2025 · 51 citations
- L4DR: LiDAR-4DRadar Fusion for Weather-Robust 3D Object DetectionXun Huang, Ziyu Xu, Hai Wu, Jinlong Wang et al.AAAI 2025 · 39 citations
- Multimodal Federated Learning via Contrastive Representation EnsembleQiying Yu, Yang Liu, Yimu Wang, Ke Xu et al.ICLR 2023 · 35 citations
Related papers
- Adaptive Hyper-graph Aggregation for Modality-Agnostic Federated LearningQ. Fan, L. ShuaiCVPR 2024
- Feature Distillation is the Better Choice for Model-Heterogeneous Federated LearningYichen Li, Xiuying Wang, Wenchao Xu, Haozhao Wang et al.NeurIPS 2025 · 6 citations
- The Best of Both Worlds: Accurate Global and Personalized Models through Federated Learning with Data-Free Hyper-Knowledge DistillationHuancheng Chen, Chianing Wang, Haris VikaloICLR 2023 · 11 citations
- Prototype-guided Bilateral Alignment Multimodal Federated LearningTianchi Liao, Lele Fu, Sheng Huang, Qing Hu et al.ICML 2026
- FedMBridge: Bridgeable Multimodal Federated LearningJiayi Chen, Aidong ZhangICML 2024 · 15 citations
