Active Finetuning: Exploiting Annotation Budget in the Pretraining-Finetuning Paradigm
Yichen Xie, Han Lu, Junchi Yan, Xiaokang Yang, Masayoshi Tomizuka, Wei Zhan
摘要
Given the large-scale data and the high annotation cost, pretraining-finetuning becomes a popular paradigm in multiple computer vision tasks. Previous research has covered both the unsupervised pretraining and supervised finetuning in this paradigm, while little attention is paid to exploiting the annotation budget for finetuning. To fill in this gap, we formally define this new active finetuning task focusing on the selection of samples for annotation in the pretrainingfinetuning paradigm. We propose a novel method called Ac-tiveFT for active finetuning task to select a subset of data distributing similarly with the entire unlabeled pool and maintaining enough diversity by optimizing a parametric model in the continuous space. We prove that the Earth Mover's distance between the distributions of the selected subset and the entire data pool is also reduced in this process. Extensive experiments show the leading performance and high efficiency of ActiveFT superior to baselines on both image classification and semantic segmentation. Our code is released at https://github.com/yichen928/ActiveFT .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper18
- Selectivity Drives Productivity: Efficient Dataset Pruning for Enhanced Transfer LearningYihua Zhang, Yimeng Zhang, Aochuan Chen, Jinghan Jia 等NeurIPS 2023 · 被引用 18 次
- Towards Free Data Selection with General-Purpose ModelsYichen Xie, Mingyu Ding, Masayoshi Tomizuka, Wei ZhanNeurIPS 2023 · 被引用 18 次
- Post-hoc Probabilistic Vision-Language ModelsAnton Baumann, Rui Li, Marcus Klasson, Santeri Mentu 等ICLR 2026 · 被引用 14 次
- Fairness without Harm: An Influence-Guided Active Sampling ApproachJinlong Pang, Jialu Wang, Zhaowei Zhu, Yuanshun Yao 等NeurIPS 2024 · 被引用 13 次
- Boundary Matters: A Bi-Level Active Finetuning MethodHan Lu, Yichen Xie, Xiaokang Yang, Junchi YanNeurIPS 2024 · 被引用 7 次
它引用的顶会 Paper19
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec 等NeurIPS 2020 · 被引用 9,171 次
- Training data-efficient image transformers & distillation through attentionHugo Touvron, Matthieu Cord, Matthijs Douze, Francisco Massa 等ICML 2021 · 被引用 8,974 次
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou 等ICCV 2021 · 被引用 8,921 次
相关 Paper
- ActiveDC: Distribution Calibration for Active FinetuningWenshuai Xu, Zhenghui Hu, Yu Lu, Jinzhou Meng 等CVPR 2024 · 被引用 5 次
- VeCAF: Vision-language Collaborative Active Finetuning with Training Objective AwarenessRongyu Zhang, Zefan Cai, Huanrui Yang, Zidong Liu 等ACM MM 2024 · 被引用 5 次
- SEPT: Towards Scalable and Efficient Visual Pre-trainingYiqi Lin, Huabin Zheng, Huaping Zhong, Jinjing Zhu 等AAAI 2023 · 被引用 2 次
- You Never Get a Second Chance To Make a Good First Impression: Seeding Active Learning for 3D Semantic SegmentationNermin Samet, Oriane Siméoni, Gilles Puy, Georgy Ponimatkin 等ICCV 2023 · 被引用 9 次
- Improving Task Diversity in Label Efficient Supervised Finetuning of LLMsAbhinav Arabelly, Jagrut Nemade, Robert D. Nowak, Jifan ZhangEMNLP 2025
