SSL: A Self-similarity Loss for Improving Generative Image Super-resolution
Du Chen, Zhengqiang Zhang, Jie Liang, Lei Zhang
摘要
Generative adversarial networks (GAN) and generative diffusion models (DM) have been widely used in real-world image super-resolution (Real-ISR) to enhance the image perceptual quality. However, these generative models are prone to generating visual artifacts and false image structures, resulting in unnatural Real-ISR results. Based on the fact that natural images exhibit high self-similarities, i.e., a local patch can have many similar patches to it in the whole image, in this work we propose a simple yet effective self-similarity loss (SSL) to improve the performance of generative Real-ISR models, enhancing the hallucination of structural and textural details while reducing the unpleasant visual artifacts. Specifically, we compute a self-similarity graph (SSG) of the ground-truth image, and enforce the SSG of Real-ISR output to be close to it. To reduce the training cost and focus on edge areas, we generate an edge mask from the ground-truth image, and compute the SSG only on the masked pixels. The proposed SSL serves as a general plug-and-play penalty, which could be easily applied to the off-the-shelf Real-ISR models. Our experiments demonstrate that, by coupling with SSL, the performance of many state-of-the-art Real-ISR models, including those GAN and DM based ones, can be largely improved, reproducing more perceptually realistic image details and eliminating many false reconstructions and visual artifacts. Codes and supplementary material are available at https://github.com/ChrisDud0257/SSL
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Generalized and Efficient 2D Gaussian Splatting for Arbitrary-Scale Super-ResolutionDu Chen, Liyi Chen, Zhengqiang Zhang, Lei ZhangICCV 2025 · 被引用 6 次
- Exploring Semantic Feature Discrimination for Perceptual Image Super-Resolution and Opinion-Unaware No-Reference Image Quality AssessmentGuanglu Dong, Xiangyu Liao, Mingyang Li, Guihuan Guo 等CVPR 2025
- Toward Generalized Image Quality Assessment: Relaxing the Perfect Reference Quality AssumptionDu Chen, Tianhe Wu, Kede Ma, Lei ZhangCVPR 2025
它引用的顶会 Paper35
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
- MUSIQ: Multi-scale Image Quality TransformerJunjie Ke, Qifei Wang, Yilin Wang, Peyman Milanfar 等ICCV 2021 · 被引用 1,325 次
相关 Paper
- Details or Artifacts: A Locally Discriminative Learning Approach to Realistic Image Super-ResolutionJie Liang, Hui Zeng, Lei ZhangCVPR 2022 · 被引用 192 次
- Human Guided Ground-Truth Generation for Realistic Image Super-ResolutionDu Chen, Jie Liang, Xindong Zhang, Ming Liu 等CVPR 2023
- Structure-Preserving Super Resolution With Gradient GuidanceCheng Ma, Yongming Rao, Yean Cheng, Ce Chen 等CVPR 2020
- Uncertainty-Aware GAN for Single Image Super ResolutionChenxi MaAAAI 2024 · 被引用 19 次
- StructSR: Refuse Spurious Details in Real-World Image Super-ResolutionYachao Li, Dong Liang, Tianyu Ding, Sheng-Jun HuangAAAI 2025 · 被引用 1 次
