CoGenSAM: Codebook-Interactive Generative Labeling for Adapting SAM to Crack Segmentation
Zhuangzhuang Chen, Nuo Chen, Dachong Li, Zhiliang Lin, Xingyu Feng, Yifan Zhang, Jie Chen, Jianqiang Li
摘要
The goal of this work is to adapt Segment Anything Models (SAM) into crack segmentation tasks via automatic label generation, thus eliminating manual annotation cost. In this regard, an intuitive approach is to extract edges of crack samples and generate labels via the dilation and erosion processes for fine-tuning SAM. However, this simple solution cannot guarantee the quality of generated labels, as crack regions will be corrupted due to the imperfect edge detection. To this end, this paper proposes CoGenSAM, a novel Codebook-interactive Generative Labeling framework that enables an annotation-free SAM fine-tuning. To achieve this, in the first stage, we pre-train a vector-quantized variational auto-encoder (VQVAE) by reconstructing the synthesized crack-like structures for learning crack-aware priors within the codebook. In the second stage, these priors help another VQVAE serve as the restoration model to restore the randomly corrupted structures into uncorrupted ones. Specifically, we propose the crack-aware contrastive-interaction to maximize the mutual information with the above priors via codebook interaction. Then, high-quality labels can be generated by restoring corrupted labels from edge detection, contributing to an annotation-free SAM fine-tuning. We collect a new dataset, Bridge2025, to address the limited availability of related bridge-oriented benchmarks. Experiments show that our performance is close to fully-supervised methods.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper17
- CrackFormer: Transformer Network for Fine-Grained Crack DetectionHuajun Liu, Xiangyu Miao, Christoph Mertz, Chengzhong Xu 等ICCV 2021 · 被引用 195 次
- Reference-Guided Pseudo-Label Generation for Medical Semantic SegmentationConstantin Marc Seibold, Simon Reiß, Jens Kleesiek, Rainer StiefelhagenAAAI 2022 · 被引用 84 次
- DeMT: Deformable Mixer Transformer for Multi-Task Learning of Dense PredictionYangyang Xu, Yibo Yang, Lefei ZhangAAAI 2023 · 被引用 81 次
- Self-Supervised Vessel Segmentation via Adversarial LearningYuxin Ma, Yang Hua, Hanming Deng, Tao Song 等ICCV 2021 · 被引用 72 次
- Joint Topology-preserving and Feature-refinement Network for Curvilinear Structure SegmentationMingfei Cheng, Kaili Zhao, Xuhong Guo, Yajing Xu 等ICCV 2021 · 被引用 54 次
相关 Paper
- CaPro: Curvilinear-aware Prompt Learning with Single Unlabeled Image for Cost-effective Curvilinear Structure SegmentationZhuangzhuang Chen, Qiangyu Chen, Chubin Ou, Xiaomeng LiAAAI 2026
- XSeg: A Large-scale X-ray Contraband Segmentation Benchmark For Real-World Security ScreeningHongxia Gao, Yixin Chen, Jiali Wen, Litao Li 等CVPR 2026 · 被引用 1 次
- Unsupervised Continual Anomaly Detection with Contrastively-Learned PromptJiaqi Liu, Kai Wu, Qiang Nie, Ying Chen 等AAAI 2024 · 被引用 54 次
- Flaws can be Applause: Unleashing Potential of Segmenting Ambiguous Objects in SAMChenxin Li, Yuzhi Huang, Wuyang Li, Hengyu Liu 等NeurIPS 2024 · 被引用 47 次
- MaskSAM: Auto-Prompt SAM with Mask Classification for Volumetric Medical Image SegmentationBin Xie, Hao Tang, Bin Duan, Dawen Cai 等ICCV 2025 · 被引用 7 次
