TUDSR: Twice Upsampling-Diffusion for Higher Super-Resolution
Zhiqiang Wu, Yitong Dong, Xian Wei
Abstract
Diffusion-based generative models have achieved remarkable success in real-world image super-resolution (SR). With tiled diffusion techniques, these models can produce high-resolution images that exceed their native-supported resolution. However, the quality of such high-resolution (e.g ) outputs often remains extremely poor, primarily due to two factors we consider: the image upsampling ratio (e.g ) exceeding the model's native-supported upsampling ratio (e.g ), and the model's native-supported resolution. In practice, training a native high-resolution model requires larger architectures, which incur significant computational overhead and GPU memory costs, making it hard on limited-resource equipment. Thus, we present TUDSR, a Twice Upsampling-Diffusion framework for higher SR. The TUDSR framework mainly consists of two stages: the first involves training at -resolution, and the second introduces a looped chunk-based training strategy at -resolution. Each stage adapts a one-step GAN architecture comprising a generator and a discriminator. Based on SD2.1-base, we develop TUDSR-S, which achieves state-of-the-art performance across multiple benchmarks. Extensive experiments further demonstrate that TUDSR-S generates high-quality images at the resolutions of and even , significantly outperforming existing approaches. Code is available at https://github.com/wuer5/TUDSR.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on23
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu et al.ICLR 2022 · 18,833 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 6,759 citations
Related papers
- TurboVSR: Fantastic Video Upscalers and Where to Find ThemZhongdao Wang, Guodongfang Zhao, Jingjing Ren, Bailan Feng et al.ICCV 2025
- FiDeSR: High-Fidelity and Detail-Preserving One-Step Diffusion Super-ResolutionAro Kim, Myeongjin Jang, Chaewon Moon, Youngjin Shin et al.CVPR 2026 · 3 citations
- Unleashing the Power of One-Step Diffusion based Image Super-Resolution via a Large-Scale Diffusion DiscriminatorJianze Li, Jiezhang Cao, Zichen Zou, Xiongfei Su et al.NeurIPS 2025 · 18 citations
- Selective Diffusion Distillation for Real-World High-Scale Image Super-ResolutionWenli Zheng, Huiyuan Fu, Zekai Xu, Xin Wang et al.AAAI 2026
- Text-Aware Real-World Image Super-Resolution via Diffusion Model with Joint Segmentation DecodersQiming Hu, Linlong Fan, Yiyan Luo, Yuhang Yu et al.NeurIPS 2025 · 7 citations
