SemiCVT: Semi-Supervised Convolutional Vision Transformer for Semantic Segmentation
Huimin Huang, Shiao Xie, Lanfen Lin, Ruofeng Tong, Yen-Wei Chen, Yuexiang Li, Hong Wang, Yawen Huang, Yefeng Zheng
Abstract
Semi-supervised learning improves data efficiency of deep models by leveraging unlabeled samples to alleviate the reliance on a large set of labeled samples. These successes concentrate on the pixel-wise consistency by using convolutional neural networks (CNNs) but fail to address both global learning capability and class-level features for unlabeled data. Recent works raise a new trend that Transformer achieves superior performance on the entire feature map in various tasks. In this paper, we unify the current dominant Mean-Teacher approaches by reconciling intramodel and inter-model properties for semi-supervised segmentation to produce a novel algorithm, SemiCVT, that absorbs the quintessence of CNNs and Transformer in a comprehensive way. Specifically, we first design a parallel CNN-Transformer architecture (CVT) with introducing an intra-model local-global interaction schema (LGI) in Fourier domain for full integration. The inter-model classwise consistency is further presented to complement the class-level statistics of CNNs and Transformer in a crossteaching manner. Extensive empirical evidence shows that SemiCVT yields consistent improvements over the state-ofthe-art methods in two public benchmarks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1b45bfa9-7e6b-40f6-9545-fb019722ee1fCited by top-tier papers8
- Earthfarsser: Versatile Spatio-Temporal Dynamical Systems Modeling in One ModelHao Wu, Yuxuan Liang, Wei Xiong, Zhengyang Zhou et al.AAAI 2024 · 58 citations
- AllSpark: Reborn Labeled Features from Unlabeled in Transformer for Semi-Supervised Semantic SegmentationHaonan Wang, Qixiang Zhang, Yi Li, Xiaomeng LiCVPR 2024 · 39 citations
- When Confidence Fails: Revisiting Pseudo-Label Selection in Semi-Supervised Semantic SegmentationPan Liu, Jinshi LiuICCV 2025 · 9 citations
- ScaleMatch: Multi-scale Consistency Enhancement for Semi-supervised Semantic SegmentationLiang Lv, Lefei ZhangAAAI 2025 · 5 citations
- Backdoor Collapse: Eliminating Unknown Threats Via Known Backdoor Aggregation In Language ModelsLiang Lin, Miao Yu, Moayad Aloqaily, Zhenhong Zhou et al.ACL 2026 · 4 citations
Builds on21
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- SegFormer: Simple and Efficient Design for Semantic Segmentation with TransformersEnze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar et al.NeurIPS 2021 · 9,661 citations
- CutMix: Regularization Strategy to Train Strong Classifiers With Localizable FeaturesSangdoo Yun, Dongyoon Han, Sanghyuk Chun, Seong Joon Oh et al.ICCV 2019 · 5,843 citations
- FixMatch: Simplifying Semi-Supervised Learning with Consistency and ConfidenceKihyuk Sohn, David Berthelot, Nicholas Carlini, Zizhao Zhang et al.NeurIPS 2020 · 5,129 citations
Related papers
- Combinatorial CNN-Transformer Learning with Manifold Constraints for Semi-supervised Medical Image SegmentationHuimin Huang, Yawen Huang, Shiao Xie, Lanfen Lin et al.AAAI 2024 · 17 citations
- Semi-Supervised Convolutional Vision Transformer with Bi-Level Uncertainty Estimation for Medical Image SegmentationHuimin Huang, Yawen Huang, Shiao Xie, Lanfen Lin et al.ACM MM 2023 · 5 citations
- Training Vision Transformers for Semi-Supervised Semantic SegmentationXinting Hu, Li Jiang, Bernt SchieleCVPR 2024
- Transformer-based Open-world Instance Segmentation with Cross-task Consistency RegularizationXizhe Xue, Dongdong Yu, Lingqiao Liu, Yu Liu et al.ACM MM 2023 · 2 citations
- Semi-supervised Medical Image Segmentation through Dual-task ConsistencyXiangde Luo, Jieneng Chen, Tao Song, Guotai WangAAAI 2021 · 754 citations
