Dynamic Neural Representational Decoders for High-Resolution Semantic Segmentation
Bowen Zhang, Yifan Liu, Zhi Tian, Chunhua Shen
Abstract
Semantic segmentation requires per-pixel prediction for a given image. Typically, the output resolution of a segmentation network is severely reduced due to the downsampling operations in the CNN backbone. Most previous methods employ upsampling decoders to recover the spatial resolution. Various decoders were designed in the literature. Here, we propose a novel decoder, termed dynamic neural representational decoder (NRD), which is simple yet significantly more efficient. As each location on the encoder's output corresponds to a local patch of the semantic labels, in this work, we represent these local patches of labels with compact neural networks. This neural representation enables our decoder to leverage the smoothness prior in the semantic label space, and thus makes our decoder more efficient. Furthermore, these neural representations are dynamically generated and conditioned on the outputs of the encoder networks. The desired semantic labels can be efficiently decoded from the neural representations, resulting in high-resolution semantic segmentation predictions. We empirically show that our proposed decoder can outperform the decoder in DeeplabV3+ with only ∼30% computational complexity, and achieve competitive performance with the methods using dilated encoders with only ∼15% computation. Experiments on the Cityscapes, ADE20K, and PASCAL Context datasets demonstrate the effectiveness and efficiency of our proposed method.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f410d8d3-cf7c-4ea6-8356-2fdfd2f8a1ddCited by top-tier papers3
- TopFormer: Token Pyramid Transformer for Mobile Semantic SegmentationWenqiang Zhang, Zilong Huang, Guozhong Luo, Tao Chen et al.CVPR 2022 · 313 citations
- SegViT: Semantic Segmentation with Plain Vision TransformersBowen Zhang, Zhi Tian, Quan Tang, Xiangxiang Chu et al.NeurIPS 2022 · 242 citations
- Boosting Semantic Segmentation from the Perspective of Explicit Class EmbeddingsYuhe Liu, Chuanjian Liu, Kai Han, Quan Tang et al.ICCV 2023 · 9 citations
Builds on4
- Implicit Neural Representations with Periodic Activation FunctionsVincent Sitzmann, Julien N. P. Martel, Alexander W. Bergman, David B. Lindell et al.NeurIPS 2020 · 4,008 citations
- CCNet: Criss-Cross Attention for Semantic SegmentationZilong Huang, Xinggang Wang, Lichao Huang, Chang Huang et al.ICCV 2019 · 2,972 citations
- Asymmetric Non-Local Neural Networks for Semantic SegmentationZhen Zhu, Mengdu Xu, Song Bai, Tengteng Huang et al.ICCV 2019 · 694 citations
- Dynamic Multi-Scale Filters for Semantic SegmentationJunjun He, Zhongying Deng, Yu QiaoICCV 2019 · 287 citations
Related papers
- HyperSeg: Patch-Wise Hypernetwork for Real-Time Semantic SegmentationYuval Nirkin, Lior Wolf, Tal HassnerCVPR 2021
- Dynamic Sampling Network for Semantic SegmentationBin Fu, Junjun He, Zhengfu Zhang, Yu QiaoAAAI 2020 · 6 citations
- SegMAN: Omni-scale Context Modeling with State Space Models and Local Attention for Semantic SegmentationYunxiang Fu, Meng Lou, Yizhou YuCVPR 2025
- Efficient Parallel Multi-Scale Detail and Semantic Encoding Network for Lightweight Semantic SegmentationXiao Liu, Xiuya Shi, Lufei Chen, Linbo Qing et al.ACM MM 2023 · 7 citations
- Rethinking Semantic Segmentation From a Sequence-to-Sequence Perspective With TransformersSixiao Zheng, Jiachen Lu, Hengshuang Zhao, Xiatian Zhu et al.CVPR 2021
