Motion Prior Knowledge Learning with Homogeneous Language Descriptions for Moving Infrared Small Target Detection
Shengjia Chen, Luping Ji, Weiwei Duan, Shuang Peng, Mao Ye
摘要
Different from traditional object detection, pure vision is not enough to infrared small target detection, due to small target size and weak background contrast. For promoting detection performance, more target representations are needed. Currently, motion representations have been proved to be one of the most potential feature kinds for infrared small target detection. Existing methods have an obvious weakness, that besides vision features, they could only capture coarse motion representations from temporal domain. With vision features, fine motion representations could be more effective to enhance detection performance. To overcome this weakness, inspired by prevalent vision-language models, we propose the first vision-language framework with motion prior knowledge learning (MoPKL). Breaking through traditional pure-vision modality, it utilizes homogeneous language descriptions, formatted for moving targets, to directionally guide vision channel learning motion prior knowledge. With the facilitation of motion-vision alignment and motion-relation mining, the motion of infrared small targets is further refined by graph attention, to generate more fine motion representations. The extensive experiments on datasets ITSDT-15K and IRDST show that our framework is effective. It could often obviously outperform other methods.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Domain-Auxiliary Infrared Moving Small Target Detection by Learning to Overlook Domain DiscrepancyShengjia Chen, Luping Ji, Shuang Peng, Sicheng Zhu 等AAAI 2026
- CodeMamba: Shifting from Target Semantics to Self-Supervised Background Manifold Learning for Singularity Detection in Infrared SequencesJingwen Ma, Xinpeng Zhang, Fan Shi, Xu Cheng 等ICML 2026
- DEFANet: Dual-Path Edge-Target Collaboration with Frequency-Aware Enhancement for Infrared Small Target DetectionShuaiyuan Du, Yang Xiao, Zhiguo CaoAAAI 2026
- Cross-domain Joint Learning with Prototype-guided Mixture-of-Experts for Infrared Moving Small Target DetectionWeiwei Duan, Luping Ji, Jianghong Huang, Sicheng Zhu 等AAAI 2026
- SeViL: Semi-supervised Vision-Language Learning with Text Prompt Guiding for Moving Infrared Small Target DetectionWeiwei Duan, Luping Ji, Jianghong Huang, Sicheng ZhuAAAI 2026
它引用的顶会 Paper5
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- FILIP: Fine-grained Interactive Language-Image Pre-TrainingLewei Yao, Runhui Huang, Lu Hou, Guansong Lu 等ICLR 2022 · 被引用 827 次
- ISNet: Shape Matters for Infrared Small Target DetectionMingjin Zhang, Rui Zhang, Yuxiang Yang, Haichen Bai 等CVPR 2022 · 被引用 556 次
- IRPruneDet: Efficient Infrared Small Target Detection via Wavelet Structure-Regularized Soft Channel PruningMingjin Zhang, Handi Yang, Jie Guo, Yunsong Li 等AAAI 2024 · 被引用 159 次
- Temporal ROI Align for Video Object RecognitionTao Gong, Kai Chen, Xinjiang Wang, Qi Chu 等AAAI 2021 · 被引用 108 次
相关 Paper
- SAIST: Segment Any Infrared Small Target Model Guided by Contrastive Language-Image PretrainingMingjin Zhang, Xiaolong Li, Fei Gao, Jie Guo 等CVPR 2025
- MOCID: Motion Context and Displacement Information Learning for Moving Infrared Small Target DetectionMingjin Zhang, Yuanjun Ouyang, Fei Gao, Jie Guo 等AAAI 2025 · 被引用 10 次
- Text-IRSTD: Leveraging Semantic Text to Promote Infrared Small Target Detection in Complex ScenesFeng Huang, Shuyuan Zheng, Zhaobing Qiu, Huanxian Liu 等ICCV 2025 · 被引用 3 次
- CHAL: Causal-guided Hierarchical Anomaly-aware Learning for Moving Infrared Small Target DetectionWeiwei Duan, Luping Ji, Shipeng Lei, Sicheng Zhu 等CVPR 2026 · 被引用 3 次
- Spatio-Temporal Context Learning with Temporal Difference Convolution for Moving Infrared Small Target DetectionHouzhang Fang, Shukai Guo, Qiuhuan Chen, Yi Chang 等AAAI 2026
