RestoreAgent: Autonomous Image Restoration Agent via Multimodal Large Language Models
Haoyu Chen, Wenbo Li, Jinjin Gu, Jingjing Ren, Sixiang Chen, Tian Ye, Renjing Pei, Kaiwen Zhou, Fenglong Song, Lei Zhu
摘要
Natural images captured by mobile devices often suffer from multiple types of degradation, such as noise, blur, and low light. Traditional image restoration methods require manual selection of specific tasks, algorithms, and execution sequences, which is time-consuming and may yield suboptimal results. All-in-one models, though capable of handling multiple tasks, typically support only a limited range and often produce overly smooth, low-fidelity outcomes due to their broad data distribution fitting. To address these challenges, we first define a new pipeline for restoring images with multiple degradations, and then introduce RestoreAgent, an intelligent image restoration system leveraging multimodal large language models. RestoreAgent autonomously assesses the type and extent of degradation in input images and performs restoration through (1) determining the appropriate restoration tasks, (2) optimizing the task sequence, (3) selecting the most suitable models, and (4) executing the restoration. Experimental results demonstrate the superior performance of RestoreAgent in handling complex degradation, surpassing human experts. Furthermore, the system's modular design facilitates the fast integration of new tasks and models, enhancing its flexibility and scalability for various applications.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper19
- 4KAgent: Agentic Any Image to 4K Super-ResolutionYushen Zuo, Qi Zheng, Mingyang Wu, Xinrui Jiang 等NeurIPS 2025 · 被引用 51 次
- Bio-Inspired Image RestorationYuning Cui, Wenqi Ren, Alois KnollNeurIPS 2025 · 被引用 21 次
- Hybrid Agents for Image RestorationBingchen Li, Xin Li, Yiting Lu, Zhibo ChenCVPR 2026 · 被引用 17 次
- FoundIR: Unleashing Million-Scale Training Data to Advance Foundation Models for Image RestorationHao Li, Xiang Chen, Jiangxin Dong, Jinhui Tang 等ICCV 2025 · 被引用 15 次
- WeatherPrompt: Multi-modality Representation Learning for All-Weather Drone Visual Geo-LocalizationJiahao Wen, Hang Yu, Zhedong ZhengNeurIPS 2025 · 被引用 11 次
它引用的顶会 Paper31
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Restormer: Efficient Transformer for High-Resolution Image RestorationSyed Waqas Zamir, Aditya Arora, Salman Khan, Munawar Hayat 等CVPR 2022 · 被引用 3,348 次
- PaLM-E: An Embodied Multimodal Language ModelDanny Driess, Fei Xia, Mehdi S. M. Sajjadi, Corey Lynch 等ICML 2023 · 被引用 2,601 次
- FFA-Net: Feature Fusion Attention Network for Single Image DehazingXu Qin, Zhilin Wang, Yuanchao Bai, Xiaodong Xie 等AAAI 2020 · 被引用 1,828 次
- Rethinking Coarse-to-Fine Approach in Single Image DeblurringSung-Jin Cho, Seo-Won Ji, Jun-Pyo Hong, Seung-Won Jung 等ICCV 2021 · 被引用 799 次
相关 Paper
- An Intelligent Agentic System for Complex Image Restoration ProblemsKaiwen Zhu, Jinjin Gu, Zhiyuan You, Yu Qiao 等ICLR 2025
- Beyond Sequential Tools: A Unified VLM Agent System for Photographic Post-Processing via Dynamic Multi-Expert FusionHonglin Xiong, Chenjie Zhu, Jianbiao Ding, Zixuan Ni 等CVPR 2026
- Restore, Assess, Repeat: A Unified Framework for Iterative Image RestorationI-Hsiang Chen, Isma Hadji, Enrique Sanchez, Adrian Bulat 等CVPR 2026 · 被引用 2 次
- EpiAgent: An Agent-Centric System for Ancient Inscription RestorationShipeng Zhu, Ang Chen, Na Nie, Pengfei Fang 等CVPR 2026 · 被引用 2 次
- ClearAIR: A Human-Visual-Perception-Inspired All-in-One Image RestorationXu Zhang, Huan Zhang, Guoli Wang, Qian Zhang 等AAAI 2026 · 被引用 6 次
