Auto-Encoded Supervision for Perceptual Image Super-Resolution
MinKyu Lee, Sangeek Hyun, Woojin Jun, Jae-Pil Heo
Abstract
This work tackles the fidelity objective in the perceptual super-resolution (SR) task. Specifically, we address the shortcomings of pixel-level L p loss (L pix ) in the GAN-based SR framework. Since L pix is known to have a trade-off relationship against perceptual quality, prior methods often multiply a small scale factor or utilize low-pass filters. However, this work shows that these circumventions fail to address the fundamental factor that induces blurring. Accordingly, we focus on two points: 1) precisely discriminating the subcomponent of L pix that contributes to blurring, and 2) only guiding based on the factor that is free from this trade-off relationship. We show that they can be achieved in a surprisingly simple manner, with an Auto-Encoder (AE) pretrained with L pix . Accordingly, we propose the Auto-Encoded Supervision for Optimal Penalization loss (L AESOP ), a novel loss function that measures distance in the AE space 1 , instead of the raw pixel space. By simply substituting L pix with L AESOP , we can provide effective reconstruction guidance without compromising perceptual quality. Designed for simplicity, our method enables easy integration into existing SR frameworks. Extensive experiments demonstrate the effectiveness of AESOP.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 7eaff5cb-2bfb-42f6-be6d-48776f6588c0Cited by top-tier papers2
- Bridging the Perception Gap in Image Super-Resolution EvaluationShaolin Su, Josep M. Rocafort, Danna Xue, David Serrano-Lozano et al.CVPR 2026 · 4 citations
- Latent Harmony: Synergistic Unified UHD Image Restoration via Latent Space Regularization and Controllable RefinementYidi Liu, Xueyang Fu, Jie Huang, Jie Xiao et al.NeurIPS 2025 · 3 citations
Builds on23
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- MUSIQ: Multi-scale Image Quality TransformerJunjie Ke, Qifei Wang, Yilin Wang, Peyman Milanfar et al.ICCV 2021 · 1,325 citations
- Exploring CLIP for Assessing the Look and Feel of ImagesJianyi Wang, Kelvin C. K. Chan, Chen Change LoyAAAI 2023 · 1,208 citations
- Designing a Practical Degradation Model for Deep Blind Image Super-ResolutionKai Zhang, Jingyun Liang, Luc Van Gool, Radu TimofteICCV 2021 · 898 citations
- Rethinking Coarse-to-Fine Approach in Single Image DeblurringSung-Jin Cho, Seo-Won Ji, Jun-Pyo Hong, Seung-Won Jung et al.ICCV 2021 · 799 citations
Related papers
- SkipDiff: Adaptive Skip Diffusion Model for High-Fidelity Perceptual Image Super-resolutionXiaotong Luo, Yuan Xie, Yanyun Qu, Yun FuAAAI 2024 · 14 citations
- Fourier Space Losses for Efficient Perceptual Image Super-ResolutionDario Fuoli, Luc Van Gool, Radu TimofteICCV 2021 · 189 citations
- SROBB: Targeted Perceptual Loss for Single Image Super-ResolutionMohammad Saeed Rad, Behzad Bozorgtabar, Urs-Viktor Marti, Max Basler et al.ICCV 2019 · 147 citations
- Perceptual-Distortion Balanced Image Super-Resolution is a Multi-Objective Optimization ProblemQiwen Zhu, Yanjie Wang, Shilv Cai, Liqun Chen et al.ACM MM 2024 · 5 citations
- Augmenting Perceptual Super-Resolution via Image Quality PredictorsFengjia Zhang, Samrudhdhi B. Rangrej, Tristan Aumentado-Armstrong, Afsaneh Fazly et al.CVPR 2025
