Efficient Visual State Space Model for Image Deblurring
Lingshun Kong, Jiangxin Dong, Jinhui Tang, Ming-Hsuan Yang, Jinshan Pan
摘要
Convolutional neural networks (CNNs) and Vision Transformers (ViTs) have achieved excellent performance in image restoration. While ViTs generally outperform CNNs by effectively capturing long-range dependencies and inputspecific characteristics, their computational complexity increases quadratically with image resolution. This limitation hampers their practical application in high-resolution image restoration. In this paper, we propose a simple yet effective visual state space model (EVSSM) for image deblurring, leveraging the benefits of state space models (SSMs) for visual data. In contrast to existing methods that employ several fixed-direction scanning for feature extraction, which significantly increases the computational cost, we develop an efficient visual scan block that applies various geometric transformations before each SSM-based module, capturing useful non-local information and maintaining high efficiency. In addition, to more effectively capture and represent local information, we propose an efficient discriminative frequency domain-based feedforward network (EDFFN), which can effectively estimate useful frequency information for latent clear image restoration. Extensive experimental results show that the proposed EVSSM performs favorably against state-of-the-art methods on benchmark datasets and real-world images.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- 4KAgent: Agentic Any Image to 4K Super-ResolutionYushen Zuo, Qi Zheng, Mingyang Wu, Xinrui Jiang 等NeurIPS 2025 · 被引用 51 次
- LoFormer: Local Frequency Transformer for Image DeblurringXintian Mao, Jiansheng Wang, Xingran Xie, Qingli Li 等ACM MM 2024 · 被引用 44 次
- Scan Clusters, Not Pixels: A Cluster-Centric Paradigm for Efficient Ultra-high-definition Image RestorationChen Wu, Ling Wang, Zhuoran Zheng, Yuning Cui 等CVPR 2026 · 被引用 9 次
- MODEM: A Morton-Order Degradation Estimation Mechanism for Adverse Weather Image RecoveryHainuo Wang, Qiming Hu, Xiaojie GuoNeurIPS 2025 · 被引用 8 次
- EVDM: Event-based Real-World Video Deblurring with MambaZhijing Sun, Senyan Xu, Kean Liu, Runze Tian 等ICCV 2025 · 被引用 6 次
它引用的顶会 Paper23
- Efficiently Modeling Long Sequences with Structured State SpacesAlbert Gu, Karan Goel, Christopher RéICLR 2022 · 被引用 3,482 次
- Restormer: Efficient Transformer for High-Resolution Image RestorationSyed Waqas Zamir, Aditya Arora, Salman Khan, Munawar Hayat 等CVPR 2022 · 被引用 3,348 次
- VMamba: Visual State Space ModelYue Liu, Yunjie Tian, Yuzhong Zhao, Hongtian Yu 等NeurIPS 2024 · 被引用 3,199 次
- Uformer: A General U-Shaped Transformer for Image RestorationZhendong Wang, Xiaodong Cun, Jianmin Bao, Wengang Zhou 等CVPR 2022 · 被引用 1,970 次
- FFA-Net: Feature Fusion Attention Network for Single Image DehazingXu Qin, Zhilin Wang, Yuanchao Bai, Xiaodong Xie 等AAAI 2020 · 被引用 1,828 次
相关 Paper
- Learning Enriched Features via Selective State Spaces Model for Efficient Image DeblurringHu Gao, Bowen Ma, Ying Zhang, Jingfan Yang 等ACM MM 2024 · 被引用 25 次
- VSSD: Vision Mamba With Non-Causal State Space DualityYuheng Shi, Mingjia Li, Minjing Dong, Chang XuICCV 2025 · 被引用 20 次
- DAMamba: Vision State Space Model with Dynamic Adaptive ScanTanzhe Li, Caoshuo Li, Jiayi Lyu, Hongjuan Pei 等NeurIPS 2025 · 被引用 24 次
- Multi-Scale VMamba: Hierarchy in Hierarchy Visual State Space ModelYuheng Shi, Minjing Dong, Chang XuNeurIPS 2024 · 被引用 129 次
- Efficient Frequency Domain-based Transformers for High-Quality Image DeblurringLingshun Kong, Jiangxin Dong, Jianjun Ge, Mingqiang Li 等CVPR 2023
