IRMamba: Pixel Difference Mamba with Layer Restoration for Infrared Small Target Detection
Mingjin Zhang, Xiaolong Li, Fei Gao, Jie Guo
Abstract
Infrared small target detection (IRSTD) focuses on identifying small targets in infrared images. Despite advancements with deep learning, challenges persist due to the IR long-range imaging mechanism, where targets are small, dim, and easily lost in noise and background clutter. Current deep learning methods struggle to suppress noise and background interference while preserving fine details, leading to missed detections and false alarms. To address these issues, we propose IRMamba, an encoder-decoder architecture featuring Pixel Difference Mamba (PDMamba) and a Layer Restoration Module (LRM). Specifically, PDMamba integrates the intensity and directional information of pixel differences between scanning positions and their central neighborhoods into the state equation of the state space model (SSM). This enhances target detail representation and suppresses background interference by capturing local 2D dependencies from a global perspective. In addition, LRM incorporates the double-depth image prior into the iterative convergence algorithm, and utilizes the inter-layer interrelationships to gradually reverse the separation of the target layer, achieving noise suppression and refined reconstruction of the image mask. Experiments conducted on multiple public datasets, including NUAA-SIRST, NUDT-SIRST, and IRSTD-1K, demonstrate the significant advantages of IRMamba over SOTA methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 25332a73-d42b-4053-b723-5dfcdc779697Builds on8
- VMamba: Visual State Space ModelYue Liu, Yunjie Tian, Yuzhong Zhao, Hongtian Yu et al.NeurIPS 2024 · 3,199 citations
- Vision Mamba: Efficient Visual Representation Learning with Bidirectional State Space ModelLianghui Zhu, Bencheng Liao, Qian Zhang, Xinlong Wang et al.ICML 2024 · 1,725 citations
- ISNet: Shape Matters for Infrared Small Target DetectionMingjin Zhang, Rui Zhang, Yuxiang Yang, Haichen Bai et al.CVPR 2022 · 556 citations
- Miss Detection vs. False Alarm: Adversarial Learning for Small Object Segmentation in Infrared ImagesHuan Wang, Luping Zhou, Lei WangICCV 2019 · 407 citations
- Deep Generalized Unfolding Networks for Image RestorationChong Mou, Qian Wang, Jian ZhangCVPR 2022 · 257 citations
Related papers
- CodeMamba: Shifting from Target Semantics to Self-Supervised Background Manifold Learning for Singularity Detection in Infrared SequencesJingwen Ma, Xinpeng Zhang, Fan Shi, Xu Cheng et al.ICML 2026
- Spatial-Frequency Mamba Collaborative Learning Network for Infrared Small Target DetectionYongji Li, Luping WangACM MM 2025 · 1 citation
- MOCID: Motion Context and Displacement Information Learning for Moving Infrared Small Target DetectionMingjin Zhang, Yuanjun Ouyang, Fei Gao, Jie Guo et al.AAAI 2025 · 10 citations
- Unleashing the Power of Generic Segmentation Model: A Simple Baseline for Infrared Small Target DetectionMingjin Zhang, Chi Zhang, Qiming Zhang, Yunsong Li et al.ACM MM 2024 · 33 citations
- Spatio-Temporal Context Learning with Temporal Difference Convolution for Moving Infrared Small Target DetectionHouzhang Fang, Shukai Guo, Qiuhuan Chen, Yi Chang et al.AAAI 2026
