CoGenSAM: Codebook-Interactive Generative Labeling for Adapting SAM to Crack Segmentation
Zhuangzhuang Chen, Nuo Chen, Dachong Li, Zhiliang Lin, Xingyu Feng, Yifan Zhang, Jie Chen, Jianqiang Li
Abstract
The goal of this work is to adapt Segment Anything Models (SAM) into crack segmentation tasks via automatic label generation, thus eliminating manual annotation cost. In this regard, an intuitive approach is to extract edges of crack samples and generate labels via the dilation and erosion processes for fine-tuning SAM. However, this simple solution cannot guarantee the quality of generated labels, as crack regions will be corrupted due to the imperfect edge detection. To this end, this paper proposes CoGenSAM, a novel Codebook-interactive Generative Labeling framework that enables an annotation-free SAM fine-tuning. To achieve this, in the first stage, we pre-train a vector-quantized variational auto-encoder (VQVAE) by reconstructing the synthesized crack-like structures for learning crack-aware priors within the codebook. In the second stage, these priors help another VQVAE serve as the restoration model to restore the randomly corrupted structures into uncorrupted ones. Specifically, we propose the crack-aware contrastive-interaction to maximize the mutual information with the above priors via codebook interaction. Then, high-quality labels can be generated by restoring corrupted labels from edge detection, contributing to an annotation-free SAM fine-tuning. We collect a new dataset, Bridge2025, to address the limited availability of related bridge-oriented benchmarks. Experiments show that our performance is close to fully-supervised methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9bfd8a1c-2d92-46f3-828f-56d9133b854aBuilds on17
- CrackFormer: Transformer Network for Fine-Grained Crack DetectionHuajun Liu, Xiangyu Miao, Christoph Mertz, Chengzhong Xu et al.ICCV 2021 · 195 citations
- Reference-Guided Pseudo-Label Generation for Medical Semantic SegmentationConstantin Marc Seibold, Simon Reiß, Jens Kleesiek, Rainer StiefelhagenAAAI 2022 · 84 citations
- DeMT: Deformable Mixer Transformer for Multi-Task Learning of Dense PredictionYangyang Xu, Yibo Yang, Lefei ZhangAAAI 2023 · 81 citations
- Self-Supervised Vessel Segmentation via Adversarial LearningYuxin Ma, Yang Hua, Hanming Deng, Tao Song et al.ICCV 2021 · 72 citations
- Joint Topology-preserving and Feature-refinement Network for Curvilinear Structure SegmentationMingfei Cheng, Kaili Zhao, Xuhong Guo, Yajing Xu et al.ICCV 2021 · 54 citations
Related papers
- CaPro: Curvilinear-aware Prompt Learning with Single Unlabeled Image for Cost-effective Curvilinear Structure SegmentationZhuangzhuang Chen, Qiangyu Chen, Chubin Ou, Xiaomeng LiAAAI 2026
- XSeg: A Large-scale X-ray Contraband Segmentation Benchmark For Real-World Security ScreeningHongxia Gao, Yixin Chen, Jiali Wen, Litao Li et al.CVPR 2026 · 1 citation
- Unsupervised Continual Anomaly Detection with Contrastively-Learned PromptJiaqi Liu, Kai Wu, Qiang Nie, Ying Chen et al.AAAI 2024 · 54 citations
- Flaws can be Applause: Unleashing Potential of Segmenting Ambiguous Objects in SAMChenxin Li, Yuzhi Huang, Wuyang Li, Hengyu Liu et al.NeurIPS 2024 · 47 citations
- MaskSAM: Auto-Prompt SAM with Mask Classification for Volumetric Medical Image SegmentationBin Xie, Hao Tang, Bin Duan, Dawen Cai et al.ICCV 2025 · 7 citations
