VQ-Seg: Vector-Quantized Token Perturbation for Semi-Supervised Medical Image Segmentation
Sicheng Yang, Zhaohu Xing, Lei Zhu
Abstract
Consistency learning with feature perturbation is a widely used strategy in semi-supervised medical image segmentation. However, many existing perturbation methods rely on dropout, and thus require a careful manual tuning of the dropout rate, which is a sensitive hyperparameter and often difficult to optimize and may lead to suboptimal regularization. To overcome this limitation, we propose VQ-Seg, the first approach to employ vector quantization (VQ) to discretize the feature space and introduce a novel and controllable Quantized Perturbation Module (QPM) that replaces dropout. Our QPM perturbs discrete representations by shuffling the spatial locations of codebook indices, enabling effective and controllable regularization. To mitigate potential information loss caused by quantization, we design a dual-branch architecture where the post-quantization feature space is shared by both image reconstruction and segmentation tasks. Moreover, we introduce a Post-VQ Feature Adapter (PFA) to incorporate guidance from a foundation model (FM), supplementing the high-level semantic information lost during quantization. Furthermore, we collect a large-scale Lung Cancer (LC) dataset comprising 828 CT scans annotated for central-type lung carcinoma. Extensive experiments on the LC dataset and other public benchmarks demonstrate the effectiveness of our method, which outperforms state-of-the-art approaches. Code available at: https://github.com/script-Yang/VQ-Seg.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 79a23e4c-4725-460d-adf0-cc8d08a54149Cited by top-tier papers1
Ask how each one uses itBuilds on12
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale PredictionKeyu Tian, Yi Jiang, Zehuan Yuan, Bingyue Peng et al.NeurIPS 2024 · 1,199 citations
- Rethinking Semi-Supervised Medical Image Segmentation: A Variance-Reduction PerspectiveChenyu You, Weicheng Dai, Yifei Min, Fenglin Liu et al.NeurIPS 2023 · 147 citations
- Scaling the Codebook Size of VQ-GAN to 100, 000 with a Utilization Rate of 99%Lei Zhu, Fangyun Wei, Yanye Lu, Dong ChenNeurIPS 2024 · 52 citations
Related papers
- GapMatch: Bridging Instance and Model Perturbations for Enhanced Semi-Supervised Medical Image SegmentationWei Huang, Lei Zhang, Zizhou Wang, Yan WangAAAI 2025 · 8 citations
- Adaptive Bidirectional Displacement for Semi-Supervised Medical Image SegmentationHanyang Chi, Jian Pang, Bingfeng Zhang, Weifeng LiuCVPR 2024
- Vector Quantization Prompting for Continual LearningLi Jiao, Qiuxia Lai, Yu Li, Qiang XuNeurIPS 2024 · 15 citations
- Regularized Vector Quantization for Tokenized Image SynthesisJiahui Zhang, Fangneng Zhan, Christian Theobalt, Shijian LuCVPR 2023
- Semi-Supervised Learning with Variational Bayesian Inference and Maximum Uncertainty RegularizationKien Do, Truyen Tran, Svetha VenkateshAAAI 2021 · 5 citations
