Scalable Diverse Model Selection for Accessible Transfer Learning
Daniel Bolya, Rohit Mittapalli, Judy Hoffman
Abstract
With the preponderance of pretrained deep learning models available off-the-shelf from model banks today, finding the best weights to fine-tune to your use-case can be a daunting task. Several methods have recently been proposed to find good models for transfer learning, but they either don't scale well to large model banks or don't perform well on the diversity of off-the-shelf models. Ideally the question we want to answer is, "given some data and a source model, can you quickly predict the model's accuracy after fine-tuning?" In this paper, we formalize this setting as "Scalable Diverse Model Selection" and propose several benchmarks for evaluating on this task. We find that existing model selection and transferability estimation methods perform poorly here and analyze why this is the case. We then introduce simple techniques to improve the performance and speed of these algorithms. Finally, we iterate on existing methods to create PARC, which outperforms all other methods on diverse model selection. We have released the benchmarks and method code † in hope to inspire future work in model selection for accessible transfer learning.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d82b9656-c7a7-4b42-884d-f38bcba28e7fCited by top-tier papers21
- Deep Model ReassemblyXingyi Yang, Daquan Zhou, Songhua Liu, Jingwen Ye et al.NeurIPS 2022 · 162 citations
- Model Spider: Learning to Rank Pre-Trained Models EfficientlyYi-Kai Zhang, Ting-Ji Huang, Yao-Xiang Ding, De-Chuan Zhan et al.NeurIPS 2023 · 57 citations
- Quick-Tune: Quickly Learning Which Pretrained Model to Finetune and HowSebastian Pineda-Arango, Fabio Ferreira, Arlind Kadra, Frank Hutter et al.ICLR 2024 · 27 citations
- Exploring Model Transferability through the Lens of Potential EnergyXiaotong Li, Zixuan Hu, Yixiao Ge, Ying Shan et al.ICCV 2023 · 15 citations
- Inner Product-based Neural Network SimilarityWei Chen, Zichen Miao, Qiang QiuNeurIPS 2023 · 14 citations
Builds on8
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Fantastic Generalization Measures and Where to Find ThemYiding Jiang, Behnam Neyshabur, Hossein Mobahi, Dilip Krishnan et al.ICLR 2020 · 705 citations
- Task2Vec: Task Embedding for Meta-LearningAlessandro Achille, Michael Lam, Rahul Tewari, Avinash Ravichandran et al.ICCV 2019 · 359 citations
- LEEP: A New Measure to Evaluate Transferability of Learned RepresentationsCuong V. Nguyen, Tal Hassner, Matthias W. Seeger, Cédric ArchambeauICML 2020 · 279 citations
- LogME: Practical Assessment of Pre-trained Models for Transfer LearningKaichao You, Yong Liu, Jianmin Wang, Mingsheng LongICML 2021 · 253 citations
Related papers
- How NOT to benchmark your SITE metric: Beyond Static Leaderboards and Towards Realistic Evaluation.Prabhant Singh, Sibylle Hess, Joaquin VanschorenICLR 2026 · 2 citations
- Foundation Model is Efficient Multimodal Multitask Model SelectorFanqing Meng, Wenqi Shao, Zhanglin Peng, Chonghe Jiang et al.NeurIPS 2023 · 26 citations
- Transferability Metrics for Selecting Source Model EnsemblesAndrea Agostinelli, Jasper R. R. Uijlings, Thomas Mensink, Vittorio FerrariCVPR 2022 · 25 citations
- Fast and Accurate Transferability Measurement by Evaluating Intra-class Feature VarianceHuiwen Xu, U KangICCV 2023 · 12 citations
- A Two-Phase Recall-and-Select Framework for Fast Model SelectionJianwei Cui, Wenhang Shi, Honglin Tao, Wei Lu et al.ICDE 2024
