ISNet: Integrate Image-Level and Semantic-Level Context for Semantic Segmentation
Zhenchao Jin, Bin Liu, Qi Chu, Nenghai Yu
Abstract
Co-occurrent visual pattern makes aggregating contextual information a common paradigm to enhance the pixel representation for semantic image segmentation. The existing approaches focus on modeling the context from the perspective of the whole image, i.e., aggregating the image-level contextual information. Despite impressive, these methods weaken the significance of the pixel representations of the same category, i.e., the semantic-level contextual information. To address this, this paper proposes to augment the pixel representations by aggregating the image-level and semantic-level contextual information, respectively. First, an image-level context module is designed to capture the contextual information for each pixel in the whole image. Second, we aggregate the representations of the same category for each pixel where the category regions are learned under the supervision of the ground-truth segmentation. Third, we compute the similarities between each pixel representation and the image-level contextual information, the semantic-level contextual information, respectively. At last, a pixel representation is augmented by weighted aggregating both the image-level contextual information and the semantic-level contextual information with the similarities as the weights. Integrating the image-level and semantic-level context allows this paper to report state-of-the-art accuracy on four benchmarks, i.e., ADE20K, LIP, COCOStuff and Cityscapes 1.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a2d296c3-499c-4c58-84a0-d8e707c65faaCited by top-tier papers12
- Rethinking Semantic Segmentation: A Prototype ViewTianfei Zhou, Wenguan Wang, Ender Konukoglu, Luc Van GoolCVPR 2022 · 353 citations
- SegViT: Semantic Segmentation with Plain Vision TransformersBowen Zhang, Zhi Tian, Quan Tang, Xiangxiang Chu et al.NeurIPS 2022 · 242 citations
- GMMSeg: Gaussian Mixture based Generative Semantic Segmentation ModelsChen Liang, Wenguan Wang, Jiaxu Miao, Yi YangNeurIPS 2022 · 185 citations
- Semantic Diffusion Network for Semantic SegmentationHaoru Tan, Sitong Wu, Jimin PiNeurIPS 2022 · 62 citations
- Coarse-to-Fine Feature Mining for Video Semantic SegmentationGuolei Sun, Yun Liu, Henghui Ding, Thomas Probst et al.CVPR 2022 · 53 citations
Builds on7
- CCNet: Criss-Cross Attention for Semantic SegmentationZilong Huang, Xinggang Wang, Lichao Huang, Chang Huang et al.ICCV 2019 · 2,972 citations
- Asymmetric Non-Local Neural Networks for Semantic SegmentationZhen Zhu, Mengdu Xu, Song Bai, Tengteng Huang et al.ICCV 2019 · 694 citations
- Expectation-Maximization Attention Networks for Semantic SegmentationXia Li, Zhisheng Zhong, Jianlong Wu, Yibo Yang et al.ICCV 2019 · 639 citations
- ACFNet: Attentional Class Feature Network for Semantic SegmentationFan Zhang, Yanqin Chen, Zhihang Li, Zhibin Hong et al.ICCV 2019 · 297 citations
- Dynamic Multi-Scale Filters for Semantic SegmentationJunjun He, Zhongying Deng, Yu QiaoICCV 2019 · 287 citations
Related papers
- Mining Contextual Information Beyond Image for Semantic SegmentationZhenchao Jin, Tao Gong, Dongdong Yu, Qi Chu et al.ICCV 2021 · 95 citations
- Exploring Cross-Image Pixel Contrast for Semantic SegmentationWenguan Wang, Tianfei Zhou, Fisher Yu, Jifeng Dai et al.ICCV 2021 · 568 citations
- Partial Class Activation Attention for Semantic SegmentationSun'ao Liu, Hongtao Xie, Hai Xu, Yongdong Zhang et al.CVPR 2022 · 47 citations
- IDRNet: Intervention-Driven Relation Network for Semantic SegmentationZhenchao Jin, Xiaowei Hu, Lingting Zhu, Luchuan Song et al.NeurIPS 2023 · 27 citations
- Region-aware Contrastive Learning for Semantic SegmentationHanzhe Hu, Jinshi Cui, Liwei WangICCV 2021 · 132 citations
