Towards Model Agnostic Federated Learning Using Knowledge Distillation
Andrei Afonin, Sai Praneeth Karimireddy
摘要
Is it possible to design an universal API for federated learning using which an ad-hoc group of data-holders (agents) collaborate with each other and perform federated learning? Such an API would necessarily need to be model-agnostic i.e. make no assumption about the model architecture being used by the agents, and also cannot rely on having representative public data at hand. Knowledge distillation (KD) is the obvious tool of choice to design such protocols. However, surprisingly, we show that most natural KD-based federated learning protocols have poor performance. To investigate this, we propose a new theoretical framework, Federated Kernel ridge regression, which can capture both model heterogeneity as well as data heterogeneity. Our analysis shows that the degradation is largely due to a fundamental limitation of knowledge distillation under data heterogeneity. We further validate our framework by analyzing and designing new protocols based on KD. Their performance on real world experiments using neural networks, though still unsatisfactory, closely matches our theoretical predictions.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper12
- DFRD: Data-Free Robustness Distillation for Heterogeneous Federated LearningKangyang Luo, Shuai Wang, Yexuan Fu, Xiang Li 等NeurIPS 2023 · 被引用 64 次
- Resource-Adaptive Federated Learning with All-In-One Neural CompositionYiqun Mei, Pengfei Guo, Mo Zhou, Vishal PatelNeurIPS 2022 · 被引用 62 次
- Accelerated Federated Learning with Decoupled Adaptive OptimizationJiayin Jin, Jiaxiang Ren, Yang Zhou, Lingjuan Lyu 等ICML 2022 · 被引用 62 次
- TCT: Convexifying Federated Learning using Bootstrapped Neural Tangent KernelsYaodong Yu, Alexander Wei, Sai Praneeth Karimireddy, Yi Ma 等NeurIPS 2022 · 被引用 38 次
- Agglomerative Federated Learning: Empowering Larger Model Training via End-Edge-Cloud CollaborationZhiyuan Wu, Sheng Sun, Yuwei Wang, Min Liu 等INFOCOM 2024 · 被引用 29 次
它引用的顶会 Paper7
- Adaptive Federated OptimizationSashank J. Reddi, Zachary Charles, Manzil Zaheer, Zachary Garrett 等ICLR 2021 · 被引用 1,917 次
- Ensemble Distillation for Robust Model Fusion in Federated LearningTao Lin, Lingjing Kong, Sebastian U. Stich, Martin JaggiNeurIPS 2020 · 被引用 1,615 次
- Model Fusion via Optimal TransportSidak Pal Singh, Martin JaggiNeurIPS 2020 · 被引用 330 次
- Self-Distillation Amplifies Regularization in Hilbert SpaceHossein Mobahi, Mehrdad Farajtabar, Peter L. BartlettNeurIPS 2020 · 被引用 298 次
- Towards Understanding Ensemble, Knowledge Distillation and Self-Distillation in Deep LearningZeyuan Allen-Zhu, Yuanzhi LiICLR 2023 · 被引用 151 次
相关 Paper
- Towards Understanding Ensemble Distillation in Federated LearningSejun Park, Kihun Hong, Ganguk HwangICML 2023 · 被引用 9 次
- FedGMKD: An Efficient Prototype Federated Learning Framework through Knowledge Distillation and Discrepancy-Aware AggregationJianqiao Zhang, Caifeng Shan, Jungong HanNeurIPS 2024 · 被引用 35 次
- FedCD: Towards Consolidated Distillation for Heterogeneous Federated LearningYichen Li, Hang Su, Huifa Li, Haolin Yang 等AAAI 2026
- FedAKD: Federated Adaptive Knowledge Distillation via Global Knowledge Calibration and DecouplingYingchao Wang, Wenqi Niu, Hanpo HouWWW 2026
- A Hierarchical Knowledge Transfer Framework for Heterogeneous Federated LearningYongheng Deng, Ju Ren, Cheng Tang, Feng Lyu 等INFOCOM 2023 · 被引用 37 次
