AllSpark: Reborn Labeled Features from Unlabeled in Transformer for Semi-Supervised Semantic Segmentation
Haonan Wang, Qixiang Zhang, Yi Li, Xiaomeng Li
Abstract
Semi-supervised semantic segmentation (SSSS) has been proposed to alleviate the burden of time-consuming pixel-level manual labeling, which leverages limited labeled data along with larger amounts of unlabeled data. Current state-of-the-art methods train the labeled data with ground truths and unlabeled data with pseudo labels. However, the two training flows are separate, which allows labeled data to dominate the training process, resulting in low-quality pseudo labels and, consequently, sub-optimal results. To alleviate this issue, we present AllSpark<sup xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">1</sup><sup xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">1</sup>The AllSpark is a powerful Cybertronian artifact in the film of Trans-formers: Revenge of the Fallen, which can be used to reborn the Trans-formers. It aligns well with our core idea., which reborns the labeled features from unlabeled ones with the channel-wise cross-attention mechanism. We further introduce a Semantic Memory along with a Channel Se-mantic Grouping strategy to ensure that unlabeled features adequately represent labeled features. The AllSpark shed new light on the architecture level designs of SSSS rather than framework level, which avoids increasingly complicated training pipeline designs. It can also be re-garded as a flexible bottleneck module that can be seam-lessly integrated into a general transformer-based segmentation model. The proposed AllSpark outperforms existing methods across all evaluation protocols on Pas-cal, Cityscapes and COCO benchmarks without bells-and-whistles. Code and model weights are available at: https://github.com/xmed-lab/AllSpark.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 7a518aba-e767-4540-ab5c-aaf4d027f5e1Cited by top-tier papers15
- AUCSeg: AUC-oriented Pixel-level Long-tail Semantic SegmentationBoyu Han, Qianqian Xu, Zhiyong Yang, Shilong Bao et al.NeurIPS 2024 · 26 citations
- When Confidence Fails: Revisiting Pseudo-Label Selection in Semi-Supervised Semantic SegmentationPan Liu, Jinshi LiuICCV 2025 · 9 citations
- Towards Robust Pseudo-Label Learning in Semantic Segmentation: An Encoding PerspectiveWangkai Li, Rui Sun, Zhaoyang Li, Tianzhu ZhangNeurIPS 2025 · 5 citations
- S5: Scalable Semi-Supervised Semantic Segmentation in Remote SensingLiang Lv, Di Wang, Jing Zhang, Lefei ZhangAAAI 2026 · 4 citations
- ConformalSAM: Unlocking the Potential of Foundational Segmentation Models in Semi-Supervised Semantic Segmentation with Conformal PredictionDanhui Chen, Ziquan Liu, Chuxi Yang, Dan Wang et al.ICCV 2025 · 3 citations
Builds on34
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- SegFormer: Simple and Efficient Design for Semantic Segmentation with TransformersEnze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar et al.NeurIPS 2021 · 9,661 citations
- FixMatch: Simplifying Semi-Supervised Learning with Consistency and ConfidenceKihyuk Sohn, David Berthelot, Nicholas Carlini, Zizhao Zhang et al.NeurIPS 2020 · 5,129 citations
- Pyramid Vision Transformer: A Versatile Backbone for Dense Prediction without ConvolutionsWenhai Wang, Enze Xie, Xiang Li, Deng-Ping Fan et al.ICCV 2021 · 4,909 citations
Related papers
- Label-Efficient Few-Shot Semantic Segmentation with Unsupervised Meta-TrainingJianwu Li, Kaiyue Shi, Guo-Sen Xie, Xiaofeng Liu et al.AAAI 2024 · 15 citations
- ST++: Make Self-trainingWork Better for Semi-supervised Semantic SegmentationLihe Yang, Wei Zhuo, Lei Qi, Yinghuan Shi et al.CVPR 2022 · 467 citations
- Training Vision Transformers for Semi-Supervised Semantic SegmentationXinting Hu, Li Jiang, Bernt SchieleCVPR 2024
- Semi-Supervised Convolutional Vision Transformer with Bi-Level Uncertainty Estimation for Medical Image SegmentationHuimin Huang, Yawen Huang, Shiao Xie, Lanfen Lin et al.ACM MM 2023 · 5 citations
- Robust Pseudo-Labeling via Decoupled Class-Aware Filtering and Dynamic Category CorrectionJianghang Lin, Yilin Lu, Chaoyang Zhu, Yunhang Shen et al.AAAI 2026
