DNF-SR: Dual-Input and Negative-Aware Feature Fine-Tuning for Real-World Image Super-Resolution
Shuhao Han, Wenjie Liao, Hayden Vance, Hang Dong, Rui Zhang, Chun-Le Guo, Chongyi Li
Abstract
Benefiting from the powerful generative priors of diffusion models, diffusion-based real-world image super-resolution (Real-ISR) methods have demonstrated impressive performance. To achieve efficient Real-ISR, several recent works have designed one-step diffusion-based models. However, unmediatedly feeding LR into a diffusion model creates a distributional gap with the model's original input. A straightforward approach to reduce the distribution gap is to introduce noise to the LR latents. However, directly adding noise inevitably corrupts the content of the LR images. In this study, we propose DNF-SR, a Dual-input and Negative-aware Feature fine-tuning method for Real-ISR. Specifically, we use a dual-input strategy that concatenates the original LR image with the noisy LR input and feeds them into a diffusion-based image editing model, ensuring both high-fidelity one-step super-resolution and improved perceptual and content consistency. Additionally, the noise present in the noisy LR input introduces randomness and diversity into the outputs. We exploit this property and propose a post-training optimization method, Negative-aware Feature Fine-Tuning (NF²T), which guides the model toward producing higher-quality results. NF 2 T classifies multiple outputs into positive and negative subsets and then defines implicit policy improvement directions in both the image and feature spaces, thereby further enhancing the stability of the optimization. Extensive experiments show that DNF-SR outperforms other methods. Code is available at
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on29
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida et al.NeurIPS 2022 · 24,707 citations
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu et al.ICLR 2022 · 18,833 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Direct Preference Optimization: Your Language Model is Secretly a Reward ModelRafael Rafailov, Archit Sharma, Eric Mitchell, Christopher D. Manning et al.NeurIPS 2023 · 10,924 citations
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li et al.NeurIPS 2022 · 8,965 citations
Related papers
- Bridging the Distribution Gap to Harness Pretrained Diffusion Priors for Super-ResolutionJoonKyu Park, Kyoung Mu LeeICLR 2026
- FiDeSR: High-Fidelity and Detail-Preserving One-Step Diffusion Super-ResolutionAro Kim, Myeongjin Jang, Chaewon Moon, Youngjin Shin et al.CVPR 2026 · 3 citations
- DP²O-SR: Direct Perceptual Preference Optimization for Real-World Image Super-ResolutionRongyuan Wu, Lingchen Sun, Zhengqiang Zhang, Shihao Wang et al.NeurIPS 2025 · 2 citations
- One-Step Diffusion Transformer for Controllable Real-World Image Super-ResolutionYushun Fang, Yuxiang Chen, Shibo Yin, Qiang Hu et al.CVPR 2026 · 9 citations
- FaithDiff: Unleashing Diffusion Priors for Faithful Image Super-resolutionJunyang Chen, Jinshan Pan, Jiangxin DongCVPR 2025
