CaPro: Curvilinear-aware Prompt Learning with Single Unlabeled Image for Cost-effective Curvilinear Structure Segmentation
Zhuangzhuang Chen, Qiangyu Chen, Chubin Ou, Xiaomeng Li
摘要
Curvilinear structure segmentation (CSS) plays a vital role in industrial applications, including medical imaging and structural health monitoring. Recently, the strong capacity of the Segment Anything Model (SAM) has inspired its downstream application in CSS tasks. To adapt SAM to CSS tasks, previous methods heavily rely on a certain number of samples and costly pixel-level annotation, which are hard to access for a new scenario. Considering this, the goal of our work is to adapt SAM in a very cost-effective setting where only a single unlabeled image is given. This is far more challenging than the typical supervised, unsupervised, or self-supervised learning manner that needs a large number of training samples. To tackle this problem, we propose a finetuning-free SAM for curvilinear structure segmentation, called curvilinear-aware prompt learning (CaPro), which aims to automatically learn visual prompts via a single unlabeled image. In the first stage, we generate extensive curvilinear structures and oriented sub-curvilinear box annotations. To increase the realism of generated curvilinear structures, we adapt these structures into real image domains via the Fourier Transform using a single real-world unlabeled image. Now, these adapted images can be used to train our oriented sub-curvilinear detector. In the second stage, we propose the curvilinear-aware discrete representation matching to filter those unreliable detection results. Afterward, these reliable detection results can be converted into informative prompts, contributing to the cost-effective SAM adaptation to CSS tasks. Experiments demonstrate the effectiveness of CaPro on medical image and crack segmentation tasks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper11
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu 等ICLR 2022 · 被引用 18,833 次
- CrackFormer: Transformer Network for Fine-Grained Crack DetectionHuajun Liu, Xiangyu Miao, Christoph Mertz, Chengzhong Xu 等ICCV 2021 · 被引用 195 次
- Enhancing Pseudo Label Quality for Semi-supervised Domain-Generalized Medical Image SegmentationHuifeng Yao, Xiaowei Hu, Xiaomeng LiAAAI 2022 · 被引用 150 次
- MuSc: Zero-Shot Industrial Anomaly Classification and Segmentation with Mutual Scoring of the Unlabeled ImagesXurui Li, Ziming Huang, Feng Xue, Yu ZhouICLR 2024 · 被引用 76 次
- Towards Generic Semi-Supervised Framework for Volumetric Medical Image SegmentationHaonan Wang, Xiaomeng LiNeurIPS 2023 · 被引用 75 次
相关 Paper
- Dual-level Adapter Boosting Prompt-free Curvilinear Structure SegmentationKai Zhu, Li Chen, Jun ChengCVPR 2026
- Unleashing the Potential of SAM for Medical Adaptation via Hierarchical DecodingZhiheng Cheng, Qingyue Wei, Hongru Zhu, Yan Wang 等CVPR 2024
- Unsupervised Continual Anomaly Detection with Contrastively-Learned PromptJiaqi Liu, Kai Wu, Qiang Nie, Ying Chen 等AAAI 2024 · 被引用 54 次
- Improving the Generalization of Segmentation Foundation Model under Distribution Shift via Weakly Supervised AdaptationHaojie Zhang, Yongyi Su, Xun Xu, Kui JiaCVPR 2024 · 被引用 26 次
- OFL-SAM2: Prompt SAM2 with Online Few-shot Learner for Efficient Medical Image SegmentationMeng Lan, Lefei Zhang, Xiaomeng LiAAAI 2026
