SSL: A Self-similarity Loss for Improving Generative Image Super-resolution
Du Chen, Zhengqiang Zhang, Jie Liang, Lei Zhang
Abstract
Generative adversarial networks (GAN) and generative diffusion models (DM) have been widely used in real-world image super-resolution (Real-ISR) to enhance the image perceptual quality. However, these generative models are prone to generating visual artifacts and false image structures, resulting in unnatural Real-ISR results. Based on the fact that natural images exhibit high self-similarities, i.e., a local patch can have many similar patches to it in the whole image, in this work we propose a simple yet effective self-similarity loss (SSL) to improve the performance of generative Real-ISR models, enhancing the hallucination of structural and textural details while reducing the unpleasant visual artifacts. Specifically, we compute a self-similarity graph (SSG) of the ground-truth image, and enforce the SSG of Real-ISR output to be close to it. To reduce the training cost and focus on edge areas, we generate an edge mask from the ground-truth image, and compute the SSG only on the masked pixels. The proposed SSL serves as a general plug-and-play penalty, which could be easily applied to the off-the-shelf Real-ISR models. Our experiments demonstrate that, by coupling with SSL, the performance of many state-of-the-art Real-ISR models, including those GAN and DM based ones, can be largely improved, reproducing more perceptually realistic image details and eliminating many false reconstructions and visual artifacts. Codes and supplementary material are available at https://github.com/ChrisDud0257/SSL
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 7d6c6fd2-6330-4e95-acde-3b631bf35421Cited by top-tier papers3
- Generalized and Efficient 2D Gaussian Splatting for Arbitrary-Scale Super-ResolutionDu Chen, Liyi Chen, Zhengqiang Zhang, Lei ZhangICCV 2025 · 6 citations
- Exploring Semantic Feature Discrimination for Perceptual Image Super-Resolution and Opinion-Unaware No-Reference Image Quality AssessmentGuanglu Dong, Xiangyu Liao, Mingyang Li, Guihuan Guo et al.CVPR 2025
- Toward Generalized Image Quality Assessment: Relaxing the Perfect Reference Quality AssumptionDu Chen, Tianhe Wu, Kede Ma, Lei ZhangCVPR 2025
Builds on35
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
- MUSIQ: Multi-scale Image Quality TransformerJunjie Ke, Qifei Wang, Yilin Wang, Peyman Milanfar et al.ICCV 2021 · 1,325 citations
Related papers
- Details or Artifacts: A Locally Discriminative Learning Approach to Realistic Image Super-ResolutionJie Liang, Hui Zeng, Lei ZhangCVPR 2022 · 192 citations
- Human Guided Ground-Truth Generation for Realistic Image Super-ResolutionDu Chen, Jie Liang, Xindong Zhang, Ming Liu et al.CVPR 2023
- Structure-Preserving Super Resolution With Gradient GuidanceCheng Ma, Yongming Rao, Yean Cheng, Ce Chen et al.CVPR 2020
- Uncertainty-Aware GAN for Single Image Super ResolutionChenxi MaAAAI 2024 · 19 citations
- StructSR: Refuse Spurious Details in Real-World Image Super-ResolutionYachao Li, Dong Liang, Tianyu Ding, Sheng-Jun HuangAAAI 2025 · 1 citation
