QMamba: On First Exploration of Vision Mamba for Image Quality Assessment
Fengbin Guan, Xin Li, Zihao Yu, Yiting Lu, Zhibo Chen
摘要
In this work, we take the first exploration of the recently popular foundation model, i.e., State Space Model/Mamba, in image quality assessment (IQA), aiming at observing and excavating the perception potential in vision Mamba. A series of works on Mamba has shown its significant potential in various fields, e.g., segmentation and classification. However, the perception capability of Mamba remains under-explored. Consequently, we propose QMamba by revisiting and adapting the Mamba model for three crucial IQA tasks, i.e., task-specific, universal, and transferable IQA, which reveals its clear advantages over existing foundational models, e.g., Swin Transformer, ViT, and CNNs, in terms of perception and computational cost. To improve the transferability of QMamba, we propose the StylePrompt tuning paradigm, where lightweight mean and variance prompts are injected to assist task-adaptive transfer learning of pre-trained QMamba for different downstream IQA tasks. Compared with existing prompt tuning strategies, our StylePrompt enables better perceptual transfer with lower computational cost. Extensive experiments on multiple synthetic, authentic IQA datasets, and cross IQA datasets demonstrate the effectiveness of our proposed QMamba.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- DR.Experts: Differential Refinement of Distortion-Aware Experts for Blind Image Quality AssessmentBohan Fu, Guanyi Qin, Fazhan Zhang, Zihao Huang 等AAAI 2026 · 被引用 1 次
- MambaIC: State Space Models for High-Performance Learned Image CompressionFanhu Zeng, Hao Tang, Yihua Shao, Siyu Chen 等CVPR 2025
- Probabilistic Prompt Adaptation for Unified Image Aesthetics and Quality AssessmentTakayuki Hara, Yuya OtsukaCVPR 2026
它引用的顶会 Paper19
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Efficiently Modeling Long Sequences with Structured State SpacesAlbert Gu, Karan Goel, Christopher RéICLR 2022 · 被引用 3,482 次
- Vision Mamba: Efficient Visual Representation Learning with Bidirectional State Space ModelLianghui Zhu, Bencheng Liao, Qian Zhang, Xinlong Wang 等ICML 2024 · 被引用 1,725 次
- HiPPO: Recurrent Memory with Optimal Polynomial ProjectionsAlbert Gu, Tri Dao, Stefano Ermon, Atri Rudra 等NeurIPS 2020 · 被引用 1,100 次
相关 Paper
- Selective Visual Prompting in Vision MambaYifeng Yao, Zichen Liu, Zhenyu Cui, Yuxin Peng 等AAAI 2025 · 被引用 13 次
- AesMamba: Universal Image Aesthetic Assessment with State Space ModelsFei Gao, Yuhao Lin, Jiaqi Shi, Maoying Qiao 等ACM MM 2024 · 被引用 12 次
- SaMam: Style-aware State Space Model for Arbitrary Image Style TransferHongda Liu, Longguang Wang, Ye Zhang, Ziru Yu 等CVPR 2025
- Mamba-Adaptor: State Space Model Adaptor for Visual RecognitionFei Xie, Jiahao Nie, Yujin Tang, Wenkang Zhang 等CVPR 2025
- QuadMamba: Learning Quadtree-based Selective Scan for Visual State Space ModelFei Xie, Weijia Zhang, Zhongdao Wang, Chao MaNeurIPS 2024 · 被引用 40 次
