Boosting Semantic Segmentation from the Perspective of Explicit Class Embeddings
Yuhe Liu, Chuanjian Liu, Kai Han, Quan Tang, Zengchang Qin
Abstract
Semantic segmentation is a computer vision task that associates a label with each pixel in an image. Modern approaches tend to introduce class embeddings into semantic segmentation for deeply utilizing category semantics, and regard supervised class masks as final predictions. In this paper, we explore the mechanism of class embeddings and have an insight that more explicit and meaningful class embeddings can be generated based on class masks purposely. Following this observation, we propose ECENet, a new segmentation paradigm, in which class embeddings are obtained and enhanced explicitly during interacting with multi-stage image features. Based on this, we revisit the traditional decoding process and explore inverted information flow between segmentation masks and class embeddings. Furthermore, to ensure the discriminability and informativity of features from backbone, we propose a Feature Reconstruction module, which combines intrinsic and diverse branches together to ensure the concurrence of diversity and redundancy in features. Experiments show that our ECENet outperforms its counterparts on the ADE20K dataset with much less computational cost and achieves new state-of-the-art results on PASCAL-Context dataset. The code will be released at https: //gitee.com/mindspore/models and https:// github.com/Carol-lyh/ECENet .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext fd728bc3-c0e4-48f7-93db-5f2b6576c302Cited by top-tier papers2
- Images as Noisy Labels: Unleashing the Potential of the Diffusion Model for Open-Vocabulary Semantic SegmentationFan Li, Xuanbin Wang, Xuan Wang, Zhaoxiang Zhang et al.ICCV 2025 · 3 citations
- SSA-Seg: Semantic and Spatial Adaptive Pixel-level Classifier for Semantic SegmentationXiaowen Ma, Zhenliang Ni, Xinghao ChenNeurIPS 2024
Builds on23
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- SegFormer: Simple and Efficient Design for Semantic Segmentation with TransformersEnze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar et al.NeurIPS 2021 · 9,661 citations
- Pyramid Vision Transformer: A Versatile Backbone for Dense Prediction without ConvolutionsWenhai Wang, Enze Xie, Xiang Li, Deng-Ping Fan et al.ICCV 2021 · 4,909 citations
- CCNet: Criss-Cross Attention for Semantic SegmentationZilong Huang, Xinggang Wang, Lichao Huang, Chang Huang et al.ICCV 2019 · 2,972 citations
- Vision Transformers for Dense PredictionRené Ranftl, Alexey Bochkovskiy, Vladlen KoltunICCV 2021 · 2,647 citations
Related papers
- Context Prior for Scene SegmentationChangqian Yu, Jingbo Wang, Changxin Gao, Gang Yu et al.CVPR 2020
- Embedded Discriminative Attention Mechanism for Weakly Supervised Semantic SegmentationTong Wu, Junshi Huang, Guangyu Gao, Xiaoming Wei et al.CVPR 2021
- Deconfound Semantic Shift and Incompleteness in Incremental Few-shot Semantic SegmentationYirui Wu, Yuhang Xia, Hao Li, Lixin Yuan et al.AAAI 2025 · 1 citation
- Incremental Few-Shot Semantic Segmentation via Embedding Adaptive-Update and Hyper-class RepresentationGuangchen Shi, Yirui Wu, Jun Liu, Shaohua Wan et al.ACM MM 2022 · 37 citations
- Context-aware Feature Generation For Zero-shot Semantic SegmentationZhangxuan Gu, Siyuan Zhou, Li Niu, Zihan Zhao et al.ACM MM 2020 · 111 citations
