SSA-Seg: Semantic and Spatial Adaptive Pixel-level Classifier for Semantic Segmentation
Xiaowen Ma, Zhenliang Ni, Xinghao Chen
Abstract
Vanilla pixel-level classifiers for semantic segmentation are based on a certain paradigm, involving the inner product of fixed prototypes obtained from the training set and pixel features in the test image. This approach, however, encounters significant limitations, , feature deviation in the semantic domain and information loss in the spatial domain. The former struggles with large intra-class variance among pixel features from different images, while the latter fails to utilize the structured information of semantic objects effectively. This leads to blurred mask boundaries as well as a deficiency of fine-grained recognition capability. In this paper, we propose a novel Semantic and Spatial Adaptive Classifier (SSA-Seg) to address the above challenges. Specifically, we employ the coarse masks obtained from the fixed prototypes as a guide to adjust the fixed prototype towards the center of the semantic and spatial domains in the test image. The adapted prototypes in semantic and spatial domains are then simultaneously considered to accomplish classification decisions. In addition, we propose an online multi-domain distillation learning strategy to improve the adaption process. Experimental results on three publicly available benchmarks show that the proposed SSA-Seg significantly improves the segmentation performance of the baseline models with only a minimal increase in computational cost. Code is available at https://github.com/xwmaxwma/SSA-Seg.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d49c68c4-7429-432b-9f9c-35ac3f5f5d2eBuilds on33
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- SegFormer: Simple and Efficient Design for Semantic Segmentation with TransformersEnze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar et al.NeurIPS 2021 · 9,661 citations
- A ConvNet for the 2020sZhuang Liu, Hanzi Mao, Chao-Yuan Wu, Christoph Feichtenhofer et al.CVPR 2022 · 6,782 citations
- CCNet: Criss-Cross Attention for Semantic SegmentationZilong Huang, Xinggang Wang, Lichao Huang, Chang Huang et al.ICCV 2019 · 2,972 citations
- Per-Pixel Classification is Not All You Need for Semantic SegmentationBowen Cheng, Alexander G. Schwing, Alexander KirillovNeurIPS 2021 · 2,196 citations
Related papers
- Prototypical Pseudo Label Denoising and Target Structure Learning for Domain Adaptive Semantic SegmentationPan Zhang, Bo Zhang, Ting Zhang, Dong Chen et al.CVPR 2021
- Exploring High-quality Target Domain Information for Unsupervised Domain Adaptive Semantic SegmentationJunjie Li, Zilei Wang, Yuan Gao, Xiaoming HuACM MM 2022 · 23 citations
- Diving Segmentation Model into PixelsChen Gan, Zihao Yin, Kelei He, Yang Gao et al.ICLR 2024
- Exploring High-Correlation Source Domain Information for Multi-Source Domain Adaptation in Semantic SegmentationYuxiang Cai, Meng Xi, Yongheng Shang, Jianwei YinACM MM 2023 · 3 citations
- BAPA-Net: Boundary Adaptation and Prototype Alignment for Cross-domain Semantic SegmentationYahao Liu, Jinhong Deng, Xinchen Gao, Wen Li et al.ICCV 2021 · 91 citations
