Transferability Estimation using Bhattacharyya Class Separability
Michal Pándy, Andrea Agostinelli, Jasper R. R. Uijlings, Vittorio Ferrari, Thomas Mensink
Abstract
Transfer learning has become a popular method for leveraging pre-trained models in computer vision. However, without performing computationally expensive fine-tuning, it is difficult to quantify which pre-trained source models are suitable for a specific target task, or, conversely, to which tasks a pre-trained source model can be easily adapted to. In this work, we propose Gaussian Bhattacharyya Coefficient (GBC), a novel method for quantifying transferability between a source model and a target dataset. In a first step we embed all target images in the feature space defined by the source model, and represent them with per-class Gaussians. Then, we estimate their pairwise class separability using the Bhattacharyya coefficient, yielding a simple and effective measure of how well the source model transfers to the target task. We evaluate GBC on image classification tasks in the context of dataset and architecture selection. Further, we also perform experiments on the more complex semantic segmentation transferability estimation task. We demonstrate that GBC outperforms state-of-the-art transferability metrics on most evaluation criteria in the semantic segmentation settings, matches the performance of top methods for dataset transferability in image classification, and performs best on architecture selection problems for image classification.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e7695f44-09d5-4572-932f-613d721410e9Cited by top-tier papers25
- Model Spider: Learning to Rank Pre-Trained Models EfficientlyYi-Kai Zhang, Ting-Ji Huang, Yao-Xiang Ding, De-Chuan Zhan et al.NeurIPS 2023 · 57 citations
- How Far Pre-trained Models Are from Neural Collapse on the Target Dataset Informs their TransferabilityZijian Wang, Yadan Luo, Liang Zheng, Zi Huang et al.ICCV 2023 · 33 citations
- Selecting Large Language Model to Fine-tune via Rectified Scaling LawHaowei Lin, Baizhou Huang, Haotian Ye, Qinyu Chen et al.ICML 2024 · 32 citations
- Foundation Model is Efficient Multimodal Multitask Model SelectorFanqing Meng, Wenqi Shao, Zhanglin Peng, Chonghe Jiang et al.NeurIPS 2023 · 26 citations
- Capability Instruction TuningYi-Kai Zhang, De-Chuan Zhan, Han-Jia YeAAAI 2025 · 22 citations
Builds on8
- LEEP: A New Measure to Evaluate Transferability of Learned RepresentationsCuong V. Nguyen, Tal Hassner, Matthias W. Seeger, Cédric ArchambeauICML 2020 · 279 citations
- Geometric Dataset Distances via Optimal TransportDavid Alvarez-Melis, Nicolò FusiNeurIPS 2020 · 267 citations
- LogME: Practical Assessment of Pre-trained Models for Transfer LearningKaichao You, Yong Liu, Jianmin Wang, Mingsheng LongICML 2021 · 253 citations
- Transferability and Hardness of Supervised Classification TasksAnh Tuan Tran, Cuong V. Nguyen, Tal HassnerICCV 2019 · 201 citations
- MSeg: A Composite Dataset for Multi-Domain Semantic SegmentationJohn Lambert, Zhuang Liu, Ozan Sener, James Hays et al.CVPR 2020
Related papers
- Understanding the Transferability of Representations via Task-RelatednessAkshay Mehra, Yunbei Zhang, Jihun HammNeurIPS 2024 · 13 citations
- Transferability Metrics for Selecting Source Model EnsemblesAndrea Agostinelli, Jasper R. R. Uijlings, Thomas Mensink, Vittorio FerrariCVPR 2022 · 25 citations
- Building a Winning Team: Selecting Source Model Ensembles using a Submodular Transferability Estimation ApproachVimal K. B., Saketh Bachu, Tanmay Garg, Niveditha Lakshmi Narasimhan et al.ICCV 2023 · 3 citations
- Fast and Accurate Transferability Measurement by Evaluating Intra-class Feature VarianceHuiwen Xu, U KangICCV 2023 · 12 citations
- Frustratingly Easy Transferability EstimationLong-Kai Huang, Junzhou Huang, Yu Rong, Qiang Yang et al.ICML 2022 · 71 citations
