Guiding The Last Layer in Federated Learning with Pre-Trained Models
Gwen Legate, Nicolas Bernier, Lucas Page-Caccia, Edouard Oyallon, Eugene Belilovsky
Abstract
Federated Learning (FL) is an emerging paradigm that allows a model to be trained across a number of participants without sharing data. Recent works have begun to consider the effects of using pre-trained models as an initialization point for existing FL algorithms; however, these approaches ignore the vast body of efficient transfer learning literature from the centralized learning setting. Here we revisit the problem of FL from a pre-trained model considered in prior work and expand it to a set of computer vision transfer learning problems. We first observe that simply fitting a linear classification head can be efficient in many cases. We then show that in the FL setting, fitting a classifier using the Nearest Class Means (NCM) can be done exactly and orders of magnitude more efficiently than existing proposals, while obtaining strong performance. Finally, we demonstrate that using a two-stage approach of obtaining the classifier and then fine-tuning the model can yield rapid convergence and improved generalization in the federated setting. We demonstrate the potential our method has to reduce communication and compute costs while achieving better model performance. Code for our experiments is available 1 .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0ac1a730-ee55-4dc2-841b-f65ca9a555f3Cited by top-tier papers6
- Heterogeneous LoRA for Federated Fine-tuning of On-Device Foundation ModelsYae Jee Cho, Luyang Liu, Zheng Xu, Aldi Fahrezi et al.EMNLP 2024 · 36 citations
- Accelerating Heterogeneous Federated Learning with Closed-form ClassifiersEros Fanì, Raffaello Camoriano, Barbara Caputo, Marco CicconeICML 2024 · 10 citations
- Covariances for Free: Exploiting Mean Distributions for Training-free Federated LearningDipam Goswami, Simone Magistri, Kai Wang, Bartlomiej Twardowski et al.NeurIPS 2025 · 3 citations
- MaRS: Memory-Adaptive Routing for Reliable Capacity Expansion and Knowledge RetentionGang YanICLR 2026
- FedRACE: A Hierarchical and Statistical Framework for Robust Federated LearningGang Yan, Sikai Yang, Wan DuNeurIPS 2025
Builds on14
- SCAFFOLD: Stochastic Controlled Averaging for Federated LearningSai Praneeth Karimireddy, Satyen Kale, Mehryar Mohri, Sashank J. Reddi et al.ICML 2020 · 3,875 citations
- Adaptive Federated OptimizationSashank J. Reddi, Zachary Charles, Manzil Zaheer, Zachary Garrett et al.ICLR 2021 · 1,917 citations
- Personalized Federated Learning with Theoretical Guarantees: A Model-Agnostic Meta-Learning ApproachAlireza Fallah, Aryan Mokhtari, Asuman E. OzdaglarNeurIPS 2020 · 1,354 citations
- Rethinking ImageNet Pre-TrainingKaiming He, Ross B. Girshick, Piotr DollárICCV 2019 · 1,188 citations
- Fine-Tuning can Distort Pretrained Features and Underperform Out-of-DistributionAnanya Kumar, Aditi Raghunathan, Robbie Matthew Jones, Tengyu Ma et al.ICLR 2022 · 911 citations
Related papers
- Federated Learning from Pre-Trained Models: A Contrastive Learning ApproachYue Tan, Guodong Long, Jie Ma, Lu Liu et al.NeurIPS 2022 · 316 citations
- Local Superior Soups: A Catalyst for Model Merging in Cross-Silo Federated LearningMinghui Chen, Meirui Jiang, Xin Zhang, Qi Dou et al.NeurIPS 2024 · 9 citations
- Rethinking the Starting Point: Collaborative Pre-Training for Federated Downstream TasksYun-Wei Chu, Dong-Jun Han, Seyyedali Hosseinalipour, Christopher G. BrintonAAAI 2025 · 1 citation
- FiT: Parameter Efficient Few-shot Transfer Learning for Personalized and Federated Image ClassificationAliaksandra Shysheya, John Bronskill, Massimiliano Patacchiola, Sebastian Nowozin et al.ICLR 2023 · 9 citations
- AFL: A Single-Round Analytic Approach for Federated Learning with Pre-trained ModelsRun He, Kai Tong, Di Fang, Han Sun et al.CVPR 2025
