Spatial-Temporal Graph Convolutional Network for Video-Based Person Re-Identification
Jinrui Yang, Wei-Shi Zheng, Qize Yang, Ying-Cong Chen, Qi Tian
Abstract
While video-based person re-identification (Re-ID) has drawn increasing attention and made great progress in recent years, it is still very challenging to effectively overcome the occlusion problem and the visual ambiguity problem for visually similar negative samples. On the other hand, we observe that different frames of a video can provide complementary information for each other, and the structural information of pedestrians can provide extra discriminative cues for appearance features. Thus, modeling the temporal relations of different frames and the spatial relations within a frame has the potential for solving the above problems. In this work, we propose a novel Spatial-Temporal Graph Convolutional Network (STGCN) to solve these problems. The STGCN includes two GCN branches, a spatial one and a temporal one. The spatial branch extracts structural information of a human body. The temporal branch mines discriminative cues from adjacent frames. By jointly optimizing these branches, our model extracts robust spatialtemporal information that is complementary with appearance information. As shown in the experiments, our model achieves state-of-the-art results on MARS and DukeMTMC-VideoReID datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1de3a59d-0604-4137-b6cc-e366071578adCited by top-tier papers20
- Pyramid Spatial-Temporal Aggregation for Video-based Person Re-IdentificationYingquan Wang, Pingping Zhang, Shang Gao, Xia Geng et al.ICCV 2021 · 118 citations
- Video-based Person Re-identification with Spatial and Temporal Memory NetworksChanho Eom, Geon Lee, Junghyup Lee, Bumsub HamICCV 2021 · 107 citations
- Spatio-Temporal Representation Factorization for Video-based Person Re-IdentificationAbhishek Aich, Meng Zheng, Srikrishna Karanam, Terrence Chen et al.ICCV 2021 · 86 citations
- Dense Interaction Learning for Video-based Person Re-identificationTianyu He, Xin Jin, Xu Shen, Jianqiang Huang et al.ICCV 2021 · 73 citations
- Salient-to-Broad Transition for Video Person Re-identificationShutao Bai, Bingpeng Ma, Hong Chang, Rui Huang et al.CVPR 2022 · 71 citations
Builds on1
Related papers
- Keypoint Message Passing for Video-Based Person Re-identificationDi Chen, Andreas Doering, Shanshan Zhang, Jian Yang et al.AAAI 2022 · 25 citations
- ASTA-Net: Adaptive Spatio-Temporal Attention Network for Person Re-Identification in VideosXierong Zhu, Jiawei Liu, Haoze Wu, Meng Wang et al.ACM MM 2020 · 10 citations
- Spatial-Temporal Correlation and Topology Learning for Person Re-Identification in VideosJiawei Liu, Zheng-Jun Zha, Wei Wu, Kecheng Zheng et al.CVPR 2021
- Learning Multi-Granular Hypergraphs for Video-Based Person Re-IdentificationYichao Yan, Jie Qin, Jiaxin Chen, Li Liu et al.CVPR 2020
- Discriminative Spatial Feature Learning for Person Re-IdentificationPeixi Peng, Yonghong Tian, Yangru Huang, Xiangqian Wang et al.ACM MM 2020 · 5 citations
