Relation-Guided Spatial Attention and Temporal Refinement for Video-Based Person Re-Identification
Xingze Li, Wengang Zhou, Yun Zhou, Houqiang Li
Abstract
Video-based person re-identification has received considerable attention in recent years due to its significant application in video surveillance. Compared with image-based person reidentification, video-based person re-identification is characterized by a much richer context, which raises the significance of identifying informative regions and fusing the temporal information across frames. In this paper, we propose two relation-guided modules to learn reinforced feature representations for effective re-identification. First, a relation-guided spatial attention (RGSA) module is designed to explore the discriminative regions globally. The weight at each position is determined by its feature as well as the relation features from other positions, revealing the dependence between local and global contents. Based on the adaptively weighted frame-level feature, then, a relation-guided temporal refinement (RGTR) module is proposed to further refine the feature representations across frames. The learned relation information via the RGTR module enables the individual frames to complement each other in an aggregation manner, leading to robust video-level feature representations. Extensive experiments on four prevalent benchmarks verify the state-of-theart performance of the proposed method.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6efcac06-2290-4795-a416-dae4709faa6eCited by top-tier papers4
- C2SLR: Consistency-enhanced Continuous Sign Language RecognitionRonglai Zuo, Brian MakCVPR 2022 · 118 citations
- Self-Emphasizing Network for Continuous Sign Language RecognitionLianyu Hu, Liqing Gao, Zekang Liu, Wei FengAAAI 2023 · 91 citations
- Vision Meets Wireless Positioning: Effective Person Re-identification with Recurrent Context PropagationYiheng Liu, Wengang Zhou, Mao Xi, Sanjing Shen et al.ACM MM 2020 · 9 citations
- Spatial-Temporal Correlation and Topology Learning for Person Re-Identification in VideosJiawei Liu, Zheng-Jun Zha, Wei Wu, Kecheng Zheng et al.CVPR 2021
Builds on1
Related papers
- Multi-Granularity Reference-Aided Attentive Feature Aggregation for Video-Based Person Re-IdentificationZhizheng Zhang, Cuiling Lan, Wenjun Zeng, Zhibo ChenCVPR 2020
- Rethinking Temporal Fusion for Video-Based Person Re-Identification on Semantic and Time AspectXinyang Jiang, Yifei Gong, Xiaowei Guo, Qize Yang et al.AAAI 2020 · 21 citations
- Watching You: Global-Guided Reciprocal Learning for Video-Based Person Re-IdentificationXuehu Liu, Pingping Zhang, Chenyang Yu, Huchuan Lu et al.CVPR 2021
- Relation-Aware Global Attention for Person Re-IdentificationZhizheng Zhang, Cuiling Lan, Wenjun Zeng, Xin Jin et al.CVPR 2020
- Frame-Guided Region-Aligned Representation for Video Person Re-IdentificationZengqun Chen, Zhiheng Zhou, Junchu Huang, Pengyu Zhang et al.AAAI 2020 · 31 citations
