Improving Image Restoration Through Removing Degradations in Textual Representations
Jingbo Lin, Zhilu Zhang, Yuxiang Wei, Dongwei Ren, Dongsheng Jiang, Qi Tian, Wangmeng Zuo
Abstract
In this paper, we introduce a new perspective for improving image restoration by removing degradation in the textual representations of a given degraded image. Intuitively, restoration is much easier on text modality than image one. For example, it can be easily conducted by removing degradation-related words while keeping the contentaware words. Hence, we combine the advantages of images in detail description and ones of text in degradation removal to perform restoration. To address the cross-modal assistance, we propose to map the degraded images into textual representations for removing the degradations, and then convert the restored textual representations into a guidance image for assisting image restoration. In particular, We ingeniously embed an image-to-text mapper and text restoration module into CLIP-equipped text-to-image models to generate the guidance. Then, we adopt a simple coarse-to-fine approach to dynamically inject multiscale information from guidance to image restoration networks. Extensive experiments are conducted on various image restoration tasks, including deblurring, dehazing, deraining, and denoising, and all-in-one image restoration. The results showcase that our method outperforms stateof-the-art ones across all these tasks. The codes and models are available at https://github.com/mrluin/ TextualDegRemoval.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers8
- Exploiting Diffusion Prior for Real-World Image Dehazing with Unpaired TrainingYunwei Lan, Zhigao Cui, Chang Liu, Jialun Peng et al.AAAI 2025 · 39 citations
- Rethinking Transformer-Based Blind-Spot Network for Self-Supervised Image DenoisingJunyi Li, Zhilu Zhang, Wangmeng ZuoAAAI 2025 · 31 citations
- Bio-Inspired Image RestorationYuning Cui, Wenqi Ren, Alois KnollNeurIPS 2025 · 21 citations
- FoundIR: Unleashing Million-Scale Training Data to Advance Foundation Models for Image RestorationHao Li, Xiang Chen, Jiangxin Dong, Jinhui Tang et al.ICCV 2025 · 15 citations
- UniRestorer: Universal Image Restoration via Adaptively Estimating Image Degradation at Proper GranularityJingbo Lin, Zhilu Zhang, Wenbo Li, Renjing Pei et al.ICLR 2026 · 8 citations
Builds on55
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
Related papers
- Controlling Vision-Language Models for Multi-Task Image RestorationZiwei Luo, Fredrik K. Gustafsson, Zheng Zhao, Jens Sjölund et al.ICLR 2024 · 111 citations
- PromptRestorer: A Prompting Image Restoration Method with Degradation PerceptionCong Wang, Jinshan Pan, Wei Wang, Jiangxin Dong et al.NeurIPS 2023 · 109 citations
- Bilevel Layer-Positioning LoRA for Real Image DehazingYan Zhang, Long Ma, Yuxin Feng, Zhe Huang et al.CVPR 2026 · 13 citations
- Multimodal Prompt Perceiver: Empower Adaptiveness, Generalizability and Fidelity for All-in-One Image RestorationYuang Ai, Huaibo Huang, Xiaoqiang Zhou, Jiexiang Wang et al.CVPR 2024
- PromptIR: Prompting for All-in-One Image RestorationVaishnav Potlapalli, Syed Waqas Zamir, Salman H. Khan, Fahad Shahbaz KhanNeurIPS 2023 · 386 citations
