A Kernel Perspective on Distillation-based Collaborative Learning
Sejun Park, Kihun Hong, Ganguk Hwang
Abstract
Over the past decade, there is a growing interest in collaborative learning that can enhance AI models of multiple parties. However, it is still challenging to enhance performance them without sharing private data and models from individual parties. One recent promising approach is to develop distillation-based algorithms that exploit unlabeled public data but the results are still unsatisfactory in both theory and practice. To tackle this problem, we rigorously analyze a representative distillation-based algorithm in the view of kernel regression. This work provides the first theoretical results to prove the (nearly) minimax optimality of the nonparametric collaborative learning algorithm that does not directly share local data or models in massively distributed statistically heterogeneous environments. Inspired by our theoretical results, we also propose a practical distillation-based collaborative learning algorithm based on neural network architecture. Our algorithm successfully bridges the gap between our theoretical assumptions and practical settings with neural networks through feature kernel matching. We simulate various regression tasks to verify our theory and demonstrate the practical feasibility of our proposed algorithm.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 44d7d420-fb3e-4f15-924c-e7dc3796fe8cCited by top-tier papers1
Ask how each one uses itBuilds on20
- SCAFFOLD: Stochastic Controlled Averaging for Federated LearningSai Praneeth Karimireddy, Satyen Kale, Mehryar Mohri, Sashank J. Reddi et al.ICML 2020 · 3,875 citations
- Ensemble Distillation for Robust Model Fusion in Federated LearningTao Lin, Lingjing Kong, Sebastian U. Stich, Martin JaggiNeurIPS 2020 · 1,615 citations
- Personalized Federated Learning with Moreau EnvelopesCanh T. Dinh, Nguyen Hoang Tran, Tuan Dung NguyenNeurIPS 2020 · 1,542 citations
- Federated Learning with Matched AveragingHongyi Wang, Mikhail Yurochkin, Yuekai Sun, Dimitris S. Papailiopoulos et al.ICLR 2020 · 1,368 citations
- Personalized Federated Learning with Theoretical Guarantees: A Model-Agnostic Meta-Learning ApproachAlireza Fallah, Aryan Mokhtari, Asuman E. OzdaglarNeurIPS 2020 · 1,354 citations
Related papers
- Towards Understanding Ensemble Distillation in Federated LearningSejun Park, Kihun Hong, Ganguk HwangICML 2023 · 9 citations
- Decentralized Learning with Multi-Headed DistillationAndrey Zhmoginov, Mark Sandler, Nolan Miller, Gus Kristiansen et al.CVPR 2023
- Towards Model Agnostic Federated Learning Using Knowledge DistillationAndrei Afonin, Sai Praneeth KarimireddyICLR 2022 · 55 citations
- Distributed Distillation for On-Device LearningIlai Bistritz, Ariana J. Mann, Nicholas BambosNeurIPS 2020 · 97 citations
- Collaborative Learning via Prediction ConsensusDongyang Fan, Celestine Mendler-Dünner, Martin JaggiNeurIPS 2023 · 11 citations
