Resource-Aware Federated Self-Supervised Learning with Global Class Representations
Mingyi Li, Xiao Zhang, Qi Wang, Tengfei Liu, Ruofan Wu, Weiqiang Wang, Fuzhen Zhuang, Hui Xiong, Dongxiao Yu
Abstract
Due to the heterogeneous architectures and class skew, the global representation models training in resource-adaptive federated self-supervised learning face with tricky challenges: deviated representation abilities and inconsistent representation spaces . In this work, we are the first to propose a multi-teacher knowledge distillation framework, namely FedMKD , to learn global representations with whole class knowledge from heterogeneous clients even under extreme class skew. Firstly, the adaptive knowledge integration mechanism is designed to learn better representations from all heterogeneous models with deviated representation abilities. Then the weighted combination of the self-supervised loss and the distillation loss can support the global model to encode all classes from clients into a unified space. Besides, the global knowledge anchored alignment module can make the local representation spaces close to the global spaces, which further improves the representation abilities of local ones. Finally, extensive experiments conducted on two datasets demonstrate the effectiveness of FedMKD which outperforms state-of-the-art baselines 4.78% under linear evaluation on average.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8c77227c-1dd1-417c-b2df-87d6cbe7bc89Cited by top-tier papers4
- C2Prompt: Class-aware Client Knowledge Interaction for Federated Continual LearningKunlun Xu, Yibo Feng, Jiangmeng Li, Yongsheng Qi et al.NeurIPS 2025 · 2 citations
- FedAFD: Multimodal Federated Learning via Adversarial Fusion and DistillationMin Tan, Junchao Ma, Yinfu FENG, Jiajun Ding et al.CVPR 2026 · 1 citation
- Towards Privacy-preserved Pre-training of Remote Sensing Foundation Models with Federated Mutual-Guidance LearningJieyi Tan, Chengwei Zhang, Bo Dang, Yansheng LiICCV 2025 · 1 citation
- Multi-order Orchestrated Curriculum Distillation for Model-Heterogeneous Federated Graph LearningFrank Wan, Xu Cheng, Run Liu, Wenke Huang et al.NeurIPS 2025
Builds on13
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec et al.NeurIPS 2020 · 9,171 citations
- Ensemble Distillation for Robust Model Fusion in Federated LearningTao Lin, Lingjing Kong, Sebastian U. Stich, Martin JaggiNeurIPS 2020 · 1,615 citations
- Divergence-aware Federated Self-Supervised LearningWeiming Zhuang, Yonggang Wen, Shuai ZhangICLR 2022 · 123 citations
- Collaborative Unsupervised Visual Representation Learning from Decentralized DataWeiming Zhuang, Xin Gan, Yonggang Wen, Shuai Zhang et al.ICCV 2021 · 121 citations
Related papers
- A Hierarchical Knowledge Transfer Framework for Heterogeneous Federated LearningYongheng Deng, Ju Ren, Cheng Tang, Feng Lyu et al.INFOCOM 2023 · 37 citations
- FedAKD: Federated Adaptive Knowledge Distillation via Global Knowledge Calibration and DecouplingYingchao Wang, Wenqi Niu, Hanpo HouWWW 2026
- The Best of Both Worlds: Accurate Global and Personalized Models through Federated Learning with Data-Free Hyper-Knowledge DistillationHuancheng Chen, Chianing Wang, Haris VikaloICLR 2023 · 11 citations
- Multi-label Self Knowledge DistillationXucong Wang, Pengkun Wang, Shurui Zhang, Miao Fang et al.AAAI 2025 · 2 citations
- Federated Learning with Label-Masking DistillationJianghu Lu, Shikun Li, Kexin Bao, Pengju Wang et al.ACM MM 2023 · 24 citations
