Beyond Image Super-Resolution for Image Recognition with Task-Driven Perceptual Loss
Jaeha Kim, Junghun Oh, Kyoung Mu Lee
摘要
In real-world scenarios, image recognition tasks, such as semantic segmentation and object detection, often pose greater challenges due to the lack of information available within low-resolution (LR) content. Image super-resolution (SR) is one of the promising solutions for addressing the challenges. However, due to the ill-posed property of SR, it is challenging for typical SR methods to restore taskrelevant high-frequency contents, which may dilute the advantage of utilizing the SR method. Therefore, in this paper, we propose Super-Resolution for Image Recognition (SR4IR) that effectively guides the generation of SR images beneficial to achieving satisfactory image recognition performance when processing LR images. The critical component of our SR4IR is the task-driven perceptual (TDP) loss that enables the SR network to acquire task-specific knowledge from a network tailored for a specific task. Moreover, we propose a cross-quality patch mix and an alternate training framework that significantly enhances the efficacy of the TDP loss by addressing potential problems when employing the TDP loss. Through extensive experiments, we demonstrate that our SR4IR achieves outstanding task performance by generating SR images useful for a specific image recognition task, including semantic segmentation, object detection, and image classification. The implementation code is available at https://github.com/JaehaKim97/SR4IR.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Exploiting Diffusion Prior for Task-Driven Image RestorationJaeha Kim, Junghun Oh, Kyoung Mu LeeICCV 2025 · 被引用 6 次
- Delving into Cascaded Instability: A Lipschitz Continuity View on Image Restoration and Object Detection SynergyQing Zhao, Weijian Deng, Pengxu Wei, ZiYi Dong 等NeurIPS 2025 · 被引用 2 次
- CLP: A Real-World Dataset of Contaminated Lens Protectors for Robust Semantic SegmentationSungyong Park, Sooyoung Choi, Hyunseo Koh, Youngjae Choi 等CVPR 2026
- OSMamba: Omnidirectional Spectral Mamba with Dual-Domain Prior Generator for Exposure CorrectionGehui Li, Bin Chen, Chen Zhao, Lei Zhang 等CVPR 2025
它引用的顶会 Paper16
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- SegFormer: Simple and Efficient Design for Semantic Segmentation with TransformersEnze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar 等NeurIPS 2021 · 被引用 9,661 次
- CutMix: Regularization Strategy to Train Strong Classifiers With Localizable FeaturesSangdoo Yun, Dongyoon Han, Sanghyuk Chun, Seong Joon Oh 等ICCV 2019 · 被引用 5,843 次
- Restormer: Efficient Transformer for High-Resolution Image RestorationSyed Waqas Zamir, Aditya Arora, Salman Khan, Munawar Hayat 等CVPR 2022 · 被引用 3,348 次
相关 Paper
- Visual Recognition-Driven Image Restoration for Multiple Degradation with Intrinsic Semantics RecoveryZizheng Yang, Jie Huang, Jiahao Chang, Man Zhou 等CVPR 2023
- SROBB: Targeted Perceptual Loss for Single Image Super-ResolutionMohammad Saeed Rad, Behzad Bozorgtabar, Urs-Viktor Marti, Max Basler 等ICCV 2019 · 被引用 147 次
- Fourier Space Losses for Efficient Perceptual Image Super-ResolutionDario Fuoli, Luc Van Gool, Radu TimofteICCV 2021 · 被引用 189 次
- Perception-Oriented Single Image Super-Resolution using Optimal Objective EstimationSeung Ho Park, Young-Su Moon, Nam Ik ChoCVPR 2023
- SeD: Semantic-Aware Discriminator for Image Super-ResolutionBingchen Li, Xin Li, Hanxin Zhu, Yeying Jin 等CVPR 2024
