UARE: A Unified Vision-Language Model for Image Quality Assessment, Restoration, and Enhancement
Weiqi Li, Xuanyu Zhang, Bin Chen, Jingfen Xie, Yan Wang, Kexin Zhang, Junlin Li, Li Zhang, Jian Zhang, Shijie Zhao
摘要
Image quality assessment (IQA) and image restoration are fundamental problems in low-level vision. Although IQA and restoration are closely connected conceptually, most existing work treats them in isolation. Recent advances in unified multimodal understanding-generation models demonstrate promising results and indicate that stronger understanding can improve generative performance. This motivates a single model that unifies IQA and restoration and explicitly studies how IQA can guide restoration, a setting that remains largely underexplored yet highly valuable. In this paper, we propose UARE, to our knowledge the first Unified vision-language model for image quality Assessment, Restoration, and Enhancement. Built on pretrained unified understanding and generation models, we introduce a two-stage training framework. First, a progressive, easy-to-hard schedule expands from single-type distortions to higher-order mixed degradations, enabling UARE to handle multiple degradations. Second, we perform unified fine-tuning of quality understanding and restoration with interleaved text-image data, aligning IQA signals with restoration objectives. Through multi-task co-training, UARE leverages IQA to boost restoration and enhancement performance. Extensive experiments across IQA, restoration, and enhancement tasks demonstrate the effectiveness of UARE. The code and models will be available at https://github.com/lwq20020127/UARE.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper53
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- Scaling Rectified Flow Transformers for High-Resolution Image SynthesisPatrick Esser, Sumith Kulal, Andreas Blattmann, Rahim Entezari 等ICML 2024 · 被引用 3,620 次
- Restormer: Efficient Transformer for High-Resolution Image RestorationSyed Waqas Zamir, Aditya Arora, Salman Khan, Munawar Hayat 等CVPR 2022 · 被引用 3,348 次
- MUSIQ: Multi-scale Image Quality TransformerJunjie Ke, Qifei Wang, Yilin Wang, Peyman Milanfar 等ICCV 2021 · 被引用 1,325 次
- Toward Real-World Single Image Super-Resolution: A New Benchmark and a New ModelJianrui Cai, Hui Zeng, Hongwei Yong, Zisheng Cao 等ICCV 2019 · 被引用 713 次
相关 Paper
- Restore, Assess, Repeat: A Unified Framework for Iterative Image RestorationI-Hsiang Chen, Isma Hadji, Enrique Sanchez, Adrian Bulat 等CVPR 2026 · 被引用 2 次
- Controlling Vision-Language Models for Multi-Task Image RestorationZiwei Luo, Fredrik K. Gustafsson, Zheng Zhao, Jens Sjölund 等ICLR 2024 · 被引用 111 次
- Adapting Text-to-Image Generation with Feature Difference Instruction for Generic Image RestorationChao Wang, Hehe Fan, Huichen Yang, Sarvnaz Karimi 等CVPR 2025
- ClearAIR: A Human-Visual-Perception-Inspired All-in-One Image RestorationXu Zhang, Huan Zhang, Guoli Wang, Qian Zhang 等AAAI 2026 · 被引用 6 次
- Beyond Ground-Truth: Leveraging Image Quality Priors for Real-World Image RestorationFengyang Xiao, Peng Hu, Lei Xu, XingE Guo 等CVPR 2026 · 被引用 5 次
