BiCnet-TKS: Learning Efficient Spatial-Temporal Representation for Video Person Re-Identification
Ruibing Hou, Hong Chang, Bingpeng Ma, Rui Huang, Shiguang Shan
Abstract
In this paper, we present an efficient spatial-temporal representation for video person re-identification (reID). Firstly, we propose a Bilateral Complementary Network (BiCnet) for spatial complementarity modeling. Specifically, BiCnet contains two branches. Detail Branch processes frames at original resolution to preserve the detailed visual clues, and Context Branch with a down-sampling strategy is employed to capture long-range contexts. On each branch, BiCnet appends multiple parallel and diverse attention modules to discover divergent body parts for consecutive frames, so as to obtain an integral characteristic of target identity. Furthermore, a Temporal Kernel Selection (TKS) block is designed to capture short-term as well as long-term temporal relations by an adaptive mode. TKS can be inserted into BiCnet at any depth to construct BiCnet-TKS for spatial-temporal modeling. Experimental results on multiple benchmarks show that BiCnet-TKS outperforms state-of-the-arts with about 50% less computations. The source code is available at https://github.com/ blue-blue272/BiCnet-TKS.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 39b65725-0b05-4494-827e-68a6efd8392eCited by top-tier papers20
- Salient-to-Broad Transition for Video Person Re-identificationShutao Bai, Bingpeng Ma, Hong Chang, Rui Huang et al.CVPR 2022 · 71 citations
- TF-CLIP: Learning Text-Free CLIP for Video-Based Person Re-identificationChenyang Yu, Xuehu Liu, Yingquan Wang, Pingping Zhang et al.AAAI 2024 · 68 citations
- Adaptive Uncertainty-Based Learning for Text-Based Person RetrievalShenshen Li, Chen He, Xing Xu, Fumin Shen et al.AAAI 2024 · 59 citations
- Gait Recognition in the Wild with Multi-hop Temporal SwitchJinkai Zheng, Xinchen Liu, Xiaoyan Gu, Yaoqi Sun et al.ACM MM 2022 · 50 citations
- Hierarchical Spatio-Temporal Representation Learning for Gait RecognitionLei Wang, Bo Liu, Fangfang Liang, Bincheng WangICCV 2023 · 43 citations
Builds on8
- Random Erasing Data AugmentationZhun Zhong, Liang Zheng, Guoliang Kang, Shaozi Li et al.AAAI 2020 · 4,134 citations
- SlowFast Networks for Video RecognitionChristoph Feichtenhofer, Haoqi Fan, Jitendra Malik, Kaiming HeICCV 2019 · 4,104 citations
- Global-Local Temporal Representations for Video Person Re-IdentificationJianing Li, Shiliang Zhang, Jingdong Wang, Wen Gao et al.ICCV 2019 · 241 citations
- Co-Segmentation Inspired Attention Networks for Video-Based Person Re-IdentificationArulkumar Subramaniam, Athira M. Nambiar, Anurag MittalICCV 2019 · 120 citations
- Spatial-Temporal Graph Convolutional Network for Video-Based Person Re-IdentificationJinrui Yang, Wei-Shi Zheng, Qize Yang, Ying-Cong Chen et al.CVPR 2020
Related papers
- Viewing from Frequency Domain: A DCT-based Information Enhancement Network for Video Person Re-IdentificationLiangchen Liu, Xi Yang, Nannan Wang, Xinbo GaoACM MM 2021 · 11 citations
- ASTA-Net: Adaptive Spatio-Temporal Attention Network for Person Re-Identification in VideosXierong Zhu, Jiawei Liu, Haoze Wu, Meng Wang et al.ACM MM 2020 · 10 citations
- Spatial-Temporal Correlation and Topology Learning for Person Re-Identification in VideosJiawei Liu, Zheng-Jun Zha, Wei Wu, Kecheng Zheng et al.CVPR 2021
- Rethinking Temporal Fusion for Video-Based Person Re-Identification on Semantic and Time AspectXinyang Jiang, Yifei Gong, Xiaowei Guo, Qize Yang et al.AAAI 2020 · 21 citations
- Pyramid Spatial-Temporal Aggregation for Video-based Person Re-IdentificationYingquan Wang, Pingping Zhang, Shang Gao, Xia Geng et al.ICCV 2021 · 118 citations
