Temporal-Context Enhanced Detection of Heavily Occluded Pedestrians
Jialian Wu, Chunluan Zhou, Ming Yang, Qian Zhang, Yuan Li, Junsong Yuan
摘要
State-of-the-art pedestrian detectors have performed promisingly on non-occluded pedestrians, yet they are still confronted by heavy occlusions. Although many previous works have attempted to alleviate the pedestrian occlusion issue, most of them rest on still images. In this paper, we exploit the local temporal context of pedestrians in videos and propose a tube feature aggregation network (TFAN) aiming at enhancing pedestrian detectors against severe occlusions. Specifically, for an occluded pedestrian in the current frame, we iteratively search for its relevant counterparts along temporal axis to form a tube. Then, features from the tube are aggregated according to an adaptive weight to enhance the feature representations of the occluded pedestrian. Furthermore, we devise a temporally discriminative embedding module (TDEM) and a part-based relation module (PRM), respectively, which adapts our approach to better handle tube drifting and heavy occlusions. Extensive experiments are conducted on three datasets, Caltech, NightOwls and KAIST, showing that our proposed method is significantly effective for heavily occluded pedestrian detection. Moreover, we achieve the state-of-the-art performance on the Caltech and NightOwls datasets.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- Forest R-CNN: Large-Vocabulary Long-Tailed Object Detection and Instance SegmentationJialian Wu, Liangchen Song, Tiancai Wang, Qian Zhang 等ACM MM 2020 · 被引用 81 次
- 4D-Net for Learned Multi-Modal AlignmentA. J. Piergiovanni, Vincent Casser, Michael S. Ryoo, Anelia AngelovaICCV 2021 · 被引用 69 次
- Self-Mimic Learning for Small-scale Pedestrian DetectionJialian Wu, Chunluan Zhou, Qian Zhang, Ming Yang 等ACM MM 2020 · 被引用 67 次
- Efficient Video Instance Segmentation via Tracklet Query and ProposalJialian Wu, Sudhir Yarram, Hui Liang, Tian Lan 等CVPR 2022 · 被引用 33 次
- Robust Knowledge Transfer via Hybrid Forward on the Teacher-Student ModelLiangchen Song, Jialian Wu, Ming Yang, Qian Zhang 等AAAI 2021 · 被引用 13 次
它引用的顶会 Paper6
- Sequence Level Semantics Aggregation for Video Object DetectionHaiping Wu, Yuntao Chen, Naiyan Wang, Zhaoxiang ZhangICCV 2019 · 被引用 236 次
- Relation Distillation Networks for Video Object DetectionJiajun Deng, Yingwei Pan, Ting Yao, Wengang Zhou 等ICCV 2019 · 被引用 211 次
- Object Guided External Memory Network for Video Object DetectionHanming Deng, Yang Hua, Tao Song, Zongpu Zhang 等ICCV 2019 · 被引用 109 次
- Progressive Sparse Local Attention for Video Object DetectionChaoxu Guo, Bin Fan, Jie Gu, Qian Zhang 等ICCV 2019 · 被引用 95 次
- Leveraging Long-Range Temporal Relationships Between Proposals for Video Object DetectionMykhailo Shvets, Wei Liu, Alexander C. BergICCV 2019 · 被引用 91 次
相关 Paper
- ASTA-Net: Adaptive Spatio-Temporal Attention Network for Person Re-Identification in VideosXierong Zhu, Jiawei Liu, Haoze Wu, Meng Wang 等ACM MM 2020 · 被引用 10 次
- Learning Hierarchical Graph for Occluded Pedestrian DetectionGang Li, Jian Li, Shanshan Zhang, Jian YangACM MM 2020 · 被引用 11 次
- Temporal Context Enhanced Feature Aggregation for Video Object DetectionFei He, Naiyu Gao, Qiaozhe Li, Senyao Du 等AAAI 2020 · 被引用 40 次
- TF-Blender: Temporal Feature Blender for Video Object DetectionYiming Cui, Liqi Yan, Zhiwen Cao, Dongfang LiuICCV 2021 · 被引用 171 次
- Mask-Guided Attention Network for Occluded Pedestrian DetectionYanwei Pang, Jin Xie, Muhammad Haris Khan, Rao Muhammad Anwer 等ICCV 2019 · 被引用 216 次
