ResDiff: Combining CNN and Diffusion Model for Image Super-resolution
Shuyao Shang, Zhengyang Shan, Guangxing Liu, Lunqian Wang, Xinghua Wang, Zekai Zhang, Jinglin Zhang
Abstract
Adapting the Diffusion Probabilistic Model (DPM) for direct image super-resolution is wasteful, given that a simple Convolutional Neural Network (CNN) can recover the main low-frequency content. Therefore, we present ResDiff, a novel Diffusion Probabilistic Model based on Residual structure for Single Image Super-Resolution (SISR). ResDiff utilizes a combination of a CNN, which restores primary low-frequency components, and a DPM, which predicts the residual between the ground-truth image and the CNN predicted image. In contrast to the common diffusion-based methods that directly use LR space to guide the noise towards HR space, ResDiff utilizes the CNN’s initial prediction to direct the noise towards the residual space between HR space and CNN-predicted space, which not only accelerates the generation process but also acquires superior sample quality. Additionally, a frequency-domain-based loss function for CNN is introduced to facilitate its restoration, and a frequency-domain guided diffusion is designed for DPM on behalf of predicting high-frequency details. The extensive experiments on multiple benchmark datasets demonstrate that ResDiff outperforms previous diffusion based methods in terms of shorter model convergence time, superior generation quality, and more diverse samples.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6d20a950-0560-4fe9-a389-a6d19d63b0b1Cited by top-tier papers25
- DocDiff: Document Enhancement via Residual Diffusion ModelsZongyuan Yang, Baolin Liu, Yongping Xiong, Lan Yi et al.ACM MM 2023 · 55 citations
- Effective Diffusion Transformer Architecture for Image Super-ResolutionKun Cheng, Lei Yu, Zhijun Tu, Xiao He et al.AAAI 2025 · 26 citations
- ReFIR: Grounding Large Restoration Models with Retrieval AugmentationHang Guo, Tao Dai, Zhihao Ouyang, Taolin Zhang et al.NeurIPS 2024 · 24 citations
- Sign-IDD: Iconicity Disentangled Diffusion for Sign Language ProductionShengeng Tang, Jiayi He, Dan Guo, Yanyan Wei et al.AAAI 2025 · 23 citations
- Diffusion Prior Interpolation for Flexibility Real-World Face Super-ResolutionJiarui Yang, Tao Dai, Yufei Zhu, Naiqi Li et al.AAAI 2025 · 11 citations
Builds on12
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li et al.NeurIPS 2022 · 8,965 citations
- Scalable Diffusion Models with TransformersWilliam Peebles, Saining XieICCV 2023 · 5,568 citations
- Denoising Diffusion Restoration ModelsBahjat Kawar, Michael Elad, Stefano Ermon, Jiaming SongNeurIPS 2022 · 1,439 citations
Related papers
- HDW-SR: High-Frequency Guided Diffusion Model based on Wavelet Decomposition for Image Super-ResolutionChao Yang, Boqian Zhang, Jinghao Xu, Guang JiangCVPR 2026 · 1 citation
- Resfusion: Denoising Diffusion Probabilistic Models for Image Restoration Based on Prior Residual NoiseZhenning Shi, Haoshuai Zheng, Chen Xu, Changsheng Dong et al.NeurIPS 2024 · 55 citations
- Multiscale Structure Guided Diffusion for Image DeblurringMengwei Ren, Mauricio Delbracio, Hossein Talebi, Guido Gerig et al.ICCV 2023 · 120 citations
- Learning Frequency-aware Dynamic Network for Efficient Super-ResolutionWenbin Xie, Dehua Song, Chang Xu, Chunjing Xu et al.ICCV 2021 · 89 citations
- Residual Denoising Diffusion ModelsJiawei Liu, Qiang Wang, Huijie Fan, Yinong Wang et al.CVPR 2024 · 96 citations
