Dual Super-Resolution Learning for Semantic Segmentation
Li Wang, Dong Li, Yousong Zhu, Lu Tian, Yi Shan
Abstract
Current state-of-the-art semantic segmentation methods often apply high-resolution input to attain high performance, which brings large computation budgets and limits their applications on resource-constrained devices. In this paper, we propose a simple and flexible two-stream framework named Dual Super-Resolution Learning (DSRL) to effectively improve the segmentation accuracy without introducing extra computation costs. Specifically, the proposed method consists of three parts: Semantic Segmentation Super-Resolution (SSSR), Single Image Super-Resolution (SISR) and Feature Affinity (FA) module, which can keep high-resolution representations with low-resolution input while simultaneously reducing the model computation complexity. Moreover, it can be easily generalized to other tasks, e.g., human pose estimation. This simple yet effective method leads to strong representations and is evidenced by promising performance on both semantic segmentation and human pose estimation. Specifically, for semantic segmentation on CityScapes, we can achieve ≥2% higher mIoU with similar FLOPs, and keep the performance with 70% FLOPs. For human pose estimation, we can gain ≥2% mAP with the same FLOPs and maintain mAP with 30% fewer FLOPs. Code and models are available at https: //github.com/wanglixilinx/DSRL .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 46a6a52d-9cf4-4fe5-9c69-cd2b4c4c155aCited by top-tier papers18
- Hypercorrelation Squeeze for Few-Shot SegmenationJuhong Min, Dahyun Kang, Minsu ChoICCV 2021 · 413 citations
- ISDNet: Integrating Shallow and Deep Networks for Efficient Ultra-high Resolution SegmentationShaohua Guo, Liang Liu, Zhenye Gan, Yabiao Wang et al.CVPR 2022 · 66 citations
- Coarse-to-Fine Feature Mining for Video Semantic SegmentationGuolei Sun, Yun Liu, Henghui Ding, Thomas Probst et al.CVPR 2022 · 53 citations
- Dual Transfer Learning for Event-based End-task Prediction via Pluggable Event to Image TranslationLin Wang, Yujeong Chae, Kuk-Jin YoonICCV 2021 · 46 citations
- CoSeR: Bridging Image and Language for Cognitive Super-ResolutionHaoze Sun, Wenbo Li, Jianzhuang Liu, Haoyu Chen et al.CVPR 2024 · 43 citations
Related papers
- Discriminative Region Suppression for Weakly-Supervised Semantic SegmentationBeomyoung Kim, Sangeun Han, Junmo KimAAAI 2021 · 137 citations
- Revisiting Efficient Semantic Segmentation: Learning Offsets for Better Spatial and Class Feature AlignmentShi-Chen Zhang, Yunheng Li, Yu-Huan Wu, Qibin Hou et al.ICCV 2025 · 8 citations
- SPGNet: Semantic Prediction Guidance for Scene ParsingBowen Cheng, Liang-Chieh Chen, Yunchao Wei, Yukun Zhu et al.ICCV 2019 · 117 citations
- HRFormer: High-Resolution Vision Transformer for Dense PredictYuhui Yuan, Rao Fu, Lang Huang, Weihong Lin et al.NeurIPS 2021 · 357 citations
- Efficient Semantic Segmentation by Altering Resolutions for Compressed VideosYubin Hu, Yuze He, Yanghao Li, Jisheng Li et al.CVPR 2023
