Text-Guided Explorable Image Super-Resolution
Kanchana Vaishnavi Gandikota, Paramanand Chandramouli
摘要
In this paper, we introduce the problem of zero-shot textguided exploration of the solutions to open-domain image super-resolution. Our goal is to allow users to explore diverse, semantically accurate reconstructions that preserve data consistency with the low-resolution inputs for different large downsampling factors without explicitly training for these specific degradations. We propose two approaches for zero-shot text-guided super-resolution -i) modifying the generative process of text-to-image (T2I ) diffusion models to promote consistency with low-resolution inputs, and ii) incorporating language guidance into zero-shot diffusion-based restoration methods. We show that the proposed approaches result in diverse solutions that match the semantic meaning provided by the text prompt while preserving data consistency with the degraded inputs. We evaluate the proposed baselines for the task of extreme super-resolution and demonstrate advantages in terms of restoration quality, diversity, and explorability of solutions.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Self-Supervised Selective-Guided Diffusion Model for Old-Photo Face RestorationWenjie Li, Xiangyi Wang, Heng Guo, Guangwei Gao 等NeurIPS 2025 · 被引用 13 次
- LaSe-E2V: Towards Language-guided Semantic-aware Event-to-Video ReconstructionKanghao Chen, Hangyu Li, Jiazhou Zhou, Zeyu Wang 等NeurIPS 2024 · 被引用 9 次
- The Power of Context: How Multimodality Improves Image Super-ResolutionKangfu Mei, Hossein Talebi, Mojtaba Ardakani, Vishal M. Patel 等CVPR 2025
- OVID: Open-Vocabulary Intrusion DetectionFujun Han, Jingqi Ye, Chenglong Zhang, Peng YeICLR 2026
它引用的顶会 Paper43
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
相关 Paper
- Generative Powers of TenXiaojuan Wang, Janne Kontkanen, Brian Curless, Steven M. Seitz 等CVPR 2024 · 被引用 3 次
- Steered Diffusion: A Generalized Framework for Plug-and-Play Conditional Image SynthesisNithin Gopalakrishnan Nair, Anoop Cherian, Suhas Lohit, Ye Wang 等ICCV 2023 · 被引用 22 次
- TOSS: High-quality Text-guided Novel View Synthesis from a Single ImageYukai Shi, Jianan Wang, He Cao, Boshi Tang 等ICLR 2024 · 被引用 28 次
- Self-supervised Dynamic Heterogeneous Degradation Modeling for Unified Zero-Shot Image RestorationXiaowan Hu, Jing Yang, HeNan Liu, HuaQiu Li 等CVPR 2026
- Harmonizing Visual and Textual Embeddings for Zero-Shot Text-to-Image CustomizationYeji Song, Jimyeong Kim, Wonhark Park, Wonsik Shin 等AAAI 2025 · 被引用 6 次
