SPGNet: Semantic Prediction Guidance for Scene Parsing
Bowen Cheng, Liang-Chieh Chen, Yunchao Wei, Yukun Zhu, Zilong Huang, Jinjun Xiong, Thomas S. Huang, Wen-Mei Hwu, Honghui Shi
Abstract
Multi-scale context module and single-stage encoder-decoder structure are commonly employed for semantic segmentation. The multi-scale context module refers to the operations to aggregate feature responses from a large spatial extent, while the single-stage encoder-decoder structure encodes the high-level semantic information in the encoder path and recovers the boundary information in the decoder path. In contrast, multi-stage encoder-decoder networks have been widely used in human pose estimation and show superior performance than their single-stage counterpart. However, few efforts have been attempted to bring this effective design to semantic segmentation. In this work, we propose a Semantic Prediction Guidance (SPG) module which learns to re-weight the local features through the guidance from pixel-wise semantic prediction. We find that by carefully re-weighting features across stages, a two-stage encoder-decoder network coupled with our proposed SPG module can significantly outperform its one-stage counterpart with similar parameters and computations. Finally, we report experimental results on the semantic segmentation benchmark Cityscapes, in which our SPGNet attains 81.1% on the test set using only 'fine' annotations.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b37128c5-a50f-4c26-83f2-5b8aca79d11fCited by top-tier papers21
- CCNet: Criss-Cross Attention for Semantic SegmentationZilong Huang, Xinggang Wang, Lichao Huang, Chang Huang et al.ICCV 2019 · 2,972 citations
- Per-Pixel Classification is Not All You Need for Semantic SegmentationBowen Cheng, Alexander G. Schwing, Alexander KirillovNeurIPS 2021 · 2,196 citations
- Deep Hierarchical Semantic SegmentationLiulei Li, Tianfei Zhou, Wenguan Wang, Jianwu Li et al.CVPR 2022 · 181 citations
- Neural 3D Scene Reconstruction with the Manhattan-world AssumptionHaoyu Guo, Sida Peng, Haotong Lin, Qianqian Wang et al.CVPR 2022 · 152 citations
- Fully Attentional Network for Semantic SegmentationQi Song, Jie Li, Chenghong Li, Hao Guo et al.AAAI 2022 · 63 citations
Builds on1
Related papers
- Efficient Parallel Multi-Scale Detail and Semantic Encoding Network for Lightweight Semantic SegmentationXiao Liu, Xiuya Shi, Lufei Chen, Linbo Qing et al.ACM MM 2023 · 7 citations
- End-to-end Boundary Exploration for Weakly-supervised Semantic SegmentationJianjun Chen, Shancheng Fang, Hongtao Xie, Zheng-Jun Zha et al.ACM MM 2021 · 13 citations
- Joint Semantic Segmentation and Boundary Detection Using Iterative Pyramid ContextsMingmin Zhen, Jinglu Wang, Lei Zhou, Shiwei Li et al.CVPR 2020
- ISNet: Integrate Image-Level and Semantic-Level Context for Semantic SegmentationZhenchao Jin, Bin Liu, Qi Chu, Nenghai YuICCV 2021 · 88 citations
- Squeeze-and-Attention Networks for Semantic SegmentationZilong Zhong, Zhong Qiu Lin, Rene Bidart, Xiaodan Hu et al.CVPR 2020
