SHERPA: Fine-tuning Segment Anything Models with Task-relevant Guidance
Jingcheng Xie, Yinda Chen, Xiaoyu Liu, Haoyuan Shi, Zhiwei Xiong
摘要
Segment Anything Models (SAMs) often struggle with certain specialized tasks. A common approach is to fine-tune models with specific task labels, but this often leads to overfitting, introduces model bias and significantly degrades their generalization ability. To overcome these challenges, we propose SHERPA, a novel framework that leverages a smaller SAM to guide the finetuning of a larger SAM via task-relevant features. Specifically, we first leverage the Fisher Ratio Separation (FRS) module to separate high taskrelevant features and preserve the ability of the large SAM to perform other general tasks. Then, the Guiding Feature Extraction (GFE) module is used to extract representative guiding features from the fine-tuned small SAMs. We leverage small SAMs tailored for specific tasks (including natural image segmentation, biomedical image segmentation, and video object segmentation) as guidance and then evaluate the SHERPA scheme to fine-tune larger SAM series models. Our experiments demonstrate that SHERPA enhances the retention of generalization ability across those diverse tasks, by up to 11.1%, and improves specific task performance by up to 2.2%. Code:
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper23
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu 等ICLR 2022 · 被引用 18,833 次
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao 等ICCV 2023 · 被引用 13,211 次
- AdaptFormer: Adapting Vision Transformers for Scalable Visual RecognitionShoufa Chen, Chongjian Ge, Zhan Tong, Jiangliu Wang 等NeurIPS 2022 · 被引用 1,291 次
- Segment Anything in High QualityLei Ke, Mingqiao Ye, Martin Danelljan, Yifan Liu 等NeurIPS 2023 · 被引用 709 次
- Robust fine-tuning of zero-shot modelsMitchell Wortsman, Gabriel Ilharco, Jong Wook Kim, Mike Li 等CVPR 2022 · 被引用 364 次
相关 Paper
- Distilling Semantic Priors from SAM to Efficient Image Restoration ModelsQuan Zhang, Xiaoyu Liu, Wei Li, Hanting Chen 等CVPR 2024 · 被引用 20 次
- MedSAMix: A Training-Free Model Merging Approach for Medical Image SegmentationYanwu Yang, Guinan Su, Jiesi Hu, Francesco Sammarco 等AAAI 2026 · 被引用 3 次
- Uncertainty-aware Fine-tuning of Segmentation Foundation ModelsKangning Liu, Brian L. Price, Jason Kuen, Yifei Fan 等NeurIPS 2024 · 被引用 15 次
- Towards Fine-Grained Interactive Segmentation in Images and VideosYuan Yao, Qiushi Yang, Miaomiao Cui, Liefeng BoICCV 2025 · 被引用 2 次
- SAM-PARSER: Fine-Tuning SAM Efficiently by Parameter Space ReconstructionZelin Peng, Zhengqin Xu, Zhilin Zeng, Xiaokang Yang 等AAAI 2024 · 被引用 41 次
