Multispectral Pedestrian Detection with Sparsely Annotated Label
Chan Lee, Seungho Shin, Gyeong-Moon Park, Jung Uk Kim
Abstract
Although existing Sparsely Annotated Object Detection (SAOD) approaches have made progress in handling sparsely annotated environments in multispectral domain, where only some pedestrians are annotated, they still have the following limitations: (i) they lack considerations for improving the quality of pseudo-labels for missing annotations, and (ii) they rely on fixed ground truth annotations, which leads to learning only a limited range of pedestrian visual appearances in the multispectral domain. To address these issues, we propose a novel framework called Sparsely Annotated Multispectral Pedestrian Detection (SAMPD). For limitation (i), we introduce Multispectral Pedestrian-aware Adaptive Weight (MPAW) and Positive Pseudo-label Enhancement (PPE) module. Utilizing multispectral knowledge, these modules ensure the generation of high-quality pseudolabels and enable effective learning by increasing weights for high-quality pseudo-labels based on modality characteristics. To address limitation (ii), we propose an Adaptive Pedestrian Retrieval Augmentation (APRA) module, which adaptively incorporates pedestrian patches from ground-truth and dynamically integrates high-quality pseudo-labels with the ground-truth, facilitating a more diverse learning pool of pedestrians. Extensive experimental results demonstrate that our SAMPD significantly enhances performance in sparsely annotated environments within the multispectral domain. The code is available at https://github.com/VisualAIKHU/ SAMPD.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 116c1a4c-0c6e-4db9-833d-508f74d2a08bCited by top-tier papers1
Ask how each one uses itBuilds on8
- Weakly Aligned Cross-Modal Learning for Multispectral Pedestrian DetectionLu Zhang, Xiangyu Zhu, Xiangyu Chen, Xu Yang et al.ICCV 2019 · 209 citations
- Co-mining: Self-Supervised Learning for Sparsely Annotated Object DetectionTiancai Wang, Tong Yang, Jiale Cao, Xiangyu ZhangAAAI 2021 · 57 citations
- Towards Versatile Pedestrian Detector with Multisensory-Matching and Multispectral Recalling MemoryJung Uk Kim, Sungjune Park, Yong Man RoAAAI 2022 · 31 citations
- Attentive Alignment Network for Multispectral Pedestrian DetectionNuo Chen, Jin Xie, Jing Nie, Jiale Cao et al.ACM MM 2023 · 26 citations
- Learning a Dynamic Cross-Modal Network for Multispectral Pedestrian DetectionJin Xie, Rao Muhammad Anwer, Hisham Cholakkal, Jing Nie et al.ACM MM 2022 · 25 citations
Related papers
- As Pseudo-Label Free as Possible: Leveraging Adaptive Feature Generation for Sparsely Annotated Object DetectionShuilian Yao, Yu Liu, Qi Jia, Sihong Chen et al.AAAI 2025
- CoDTS: Enhancing Sparsely Supervised Collaborative Perception with a Dual Teacher-Student FrameworkYushan Han, Hui Zhang, Honglei Zhang, Jing Wang et al.AAAI 2025 · 3 citations
- PseDet: Revisiting the Power of Pseudo Label in Incremental Object DetectionQiuchen Wang, Zehui Chen, Chenhongyi Yang, Jiaming Liu et al.ICLR 2025
- SPWOOD: Sparse Partial Weakly-Supervised Oriented Object DetectionWei Zhang, Xiang Liu, Ningjing Liu, Mingxin Liu et al.ICLR 2026
- Sparse Fuse Dense: Towards High Quality 3D Detection with Depth CompletionXiaopei Wu, Liang Peng, Honghui Yang, Liang Xie et al.CVPR 2022 · 248 citations
