MambaIRv2: Attentive State Space Restoration
Hang Guo, Yong Guo, Yaohua Zha, Yulun Zhang, Wenbo Li, Tao Dai, Shu-Tao Xia, Yawei Li
Abstract
The Mamba-based image restoration backbones have recently demonstrated significant potential in balancing global reception and computational efficiency. However, the inherent causal modeling limitation of Mamba, where each token depends solely on its predecessors in the scanned sequence, restricts the full utilization of pixels across the image and thus presents new challenges in image restoration. In this work, we propose MambaIRv2, which equips Mamba with the non-causal modeling ability similar to ViTs to reach the attentive state space restoration model. Specifically, the proposed attentive state-space equation allows to attend beyond the scanned sequence and facilitate image unfolding with just one single scan. Moreover, we further introduce a semantic-guided neighboring mechanism to encourage interaction between distant but similar pixels. Extensive experiments show our Mam-baIRv2 outperforms SRFormer by even 0.35dB PSNR for lightweight SR even with 9.3% less parameters and suppresses HAT on classic SR by up to 0.29dB. Code is available at https://github.com/csguoh/MambaIR .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d9c6179c-9876-4352-8107-7e02ed067170Cited by top-tier papers26
- Spiking Meets Attention: Efficient Remote Sensing Image Super-Resolution with Attention Spiking Neural NetworksYi Xiao, Qiangqiang Yuan, Kui Jiang, Wenke Huang et al.NeurIPS 2025 · 25 citations
- Emulating Self-attention with Convolution for Efficient Image Super-ResolutionDongheon Lee, Seokju Yun, Youngmin RoICCV 2025 · 19 citations
- Scan Clusters, Not Pixels: A Cluster-Centric Paradigm for Efficient Ultra-high-definition Image RestorationChen Wu, Ling Wang, Zhuoran Zheng, Yuning Cui et al.CVPR 2026 · 9 citations
- Depth-Synergized Mamba Meets Memory Experts for All-Day Image Reflection SeparationSiyan Fang, Long Peng, Yuntao Wang, Ruonan Wei et al.AAAI 2026 · 5 citations
- IDF: Iterative Dynamic Filtering Networks for Generalizable Image DenoisingDongjin Kim, Jaekyun Ko, Muhammad Kashif Ali, Tae Hyun KimICCV 2025 · 4 citations
Builds on21
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Restormer: Efficient Transformer for High-Resolution Image RestorationSyed Waqas Zamir, Aditya Arora, Salman Khan, Munawar Hayat et al.CVPR 2022 · 3,348 citations
- Conditional Positional Encodings for Vision TransformersXiangxiang Chu, Zhi Tian, Bo Zhang, Xinlong Wang et al.ICLR 2023 · 406 citations
- Dual Aggregation Transformer for Image Super-ResolutionZheng Chen, Yulun Zhang, Jinjin Gu, Linghe Kong et al.ICCV 2023 · 345 citations
Related papers
- SF-Mamba: Rethinking State Space Model for VisionMasakazu Yoshimura, Teruaki Hayashi, Yuki Hoshino, Wei-Yao Wang et al.ICML 2026
- Vision Mamba: Efficient Visual Representation Learning with Bidirectional State Space ModelLianghui Zhu, Bencheng Liao, Qian Zhang, Xinlong Wang et al.ICML 2024 · 1,725 citations
- MaIR: A Locality- and Continuity-Preserving Mamba for Image RestorationBoyun Li, Haiyu Zhao, Wenxin Wang, Peng Hu et al.CVPR 2025
- MaskViM: Domain Generalized Semantic Segmentation with State Space ModelsJiahao Li, Yang Lu, Yuan Xie, Yanyun QuAAAI 2025 · 1 citation
- Polyline Path Masked Attention for Vision TransformerZhongchen Zhao, Chaodong Xiao, Hui Lin, Qi Xie et al.NeurIPS 2025 · 1 citation
