DR-Tune: Improving Fine-tuning of Pretrained Visual Models by Distribution Regularization with Semantic Calibration
Nan Zhou, Jiaxin Chen, Di Huang
摘要
The visual models pretrained on large-scale benchmarks encode general knowledge and prove effective in building more powerful representations for downstream tasks. Most existing approaches follow the fine-tuning paradigm, either by initializing or regularizing the downstream model based on the pretrained one. The former fails to retain the knowledge in the successive fine-tuning phase, thereby prone to be over-fitting, and the latter imposes strong constraints to the weights or feature maps of the downstream model without considering semantic drift, often incurring insufficient optimization. To deal with these issues, we propose a novel fine-tuning framework, namely distribution regularization with semantic calibration (DR-Tune). It employs distribution regularization by enforcing the downstream task head to decrease its classification error on the pretrained feature distribution, which prevents it from over-fitting while enabling sufficient training of downstream encoders. Furthermore, to alleviate the interference by semantic drift, we develop the semantic calibration (SC) module to align the global shape and class centers of the pretrained and downstream feature distributions. Extensive experiments on widely used image classification datasets show that DR-Tune consistently improves the performance when combing with various backbones under different pretraining strategies. Code is available at: https://github.com/ weeknan/DR-Tune .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Adaptive Random Feature Regularization on Fine-tuning Deep Neural NetworksShin'ya Yamaguchi, Sekitoshi Kanai, Kazuki Adachi, Daiki ChijiwaCVPR 2024 · 被引用 3 次
- TTE: Two Tokens Are Enough to Improve Parameter-Efficient TuningJiacheng Ruan, Mingye Xie, Jingsheng Gao, Xian Gao 等AAAI 2025 · 被引用 3 次
- Bayesian-LoRA: Probabilistic Low-Rank Adaptation of Large Language ModelsMoule Lin, Shuhao Guan, Andrea Patane, David Gregg 等ICML 2026
- CoVFT: Context-aware Visual Fine-tuning for Multimodal Large Language ModelsNan Zhou, Huiqun Wang, Yaoyan Zheng, Di HuangCVPR 2026
它引用的顶会 Paper19
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- Supervised Contrastive LearningPrannay Khosla, Piotr Teterwak, Chen Wang, Aaron Sarna 等NeurIPS 2020 · 被引用 7,049 次
- Unsupervised Learning of Visual Features by Contrasting Cluster AssignmentsMathilde Caron, Ishan Misra, Julien Mairal, Priya Goyal 等NeurIPS 2020 · 被引用 5,249 次
- BEiT: BERT Pre-Training of Image TransformersHangbo Bao, Li Dong, Songhao Piao, Furu WeiICLR 2022 · 被引用 3,632 次
- Rethinking ImageNet Pre-TrainingKaiming He, Ross B. Girshick, Piotr DollárICCV 2019 · 被引用 1,188 次
相关 Paper
- Debiased Fine-Tuning for Vision-Language Models by Prompt RegularizationBeier Zhu, Yulei Niu, Saeil Lee, Minhoe Hur 等AAAI 2023 · 被引用 34 次
- Proxy-FDA: Proxy-based Feature Distribution Alignment for Fine-tuning Vision Foundation Models without ForgettingChen Huang, Skyler Seto, Hadi Pouransari, Mehrdad Farajtabar 等ICML 2025
- Improved Visual Fine-tuning with Natural Language SupervisionJunyang Wang, Yuanhong Xu, Juhua Hu, Ming Yan 等ICCV 2023 · 被引用 11 次
- Distribution Alignment: A Unified Framework for Long-Tail Visual RecognitionSongyang Zhang, Zeming Li, Shipeng Yan, Xuming He 等CVPR 2021
- Co-Tuning for Transfer LearningKaichao You, Zhi Kou, Mingsheng Long, Jianmin WangNeurIPS 2020 · 被引用 105 次
