Hierarchical Spatio-Temporal Representation Learning for Gait Recognition
Lei Wang, Bo Liu, Fangfang Liang, Bincheng Wang
Abstract
Gait recognition is a biometric technique that identifies individuals by their unique walking styles, which is suitable for unconstrained environments and has a wide range of applications. While current methods focus on exploiting body part-based representations, they often neglect the hierarchical dependencies between local motion patterns. In this paper, we propose a hierarchical spatio-temporal representation learning (HSTL) framework for extracting gait features from coarse to fine. Our framework starts with a hierarchical clustering analysis to recover multi-level body structures from the whole body to local details. Next, an adaptive region-based motion extractor (ARME) is designed to learn region-independent motion features. The proposed HSTL then stacks multiple ARMEs in a topdown manner, with each ARME corresponding to a specific partition level of the hierarchy. An adaptive spatiotemporal pooling (ASTP) module is used to capture gait features at different levels of detail to perform hierarchical feature mapping. Finally, a frame-level temporal aggregation (FTA) module is employed to reduce redundant information in gait sequences through multi-scale temporal downsampling. Extensive experiments on CASIA-B, OUMVLP, GREW, and Gait3D datasets demonstrate that our method outperforms the state-of-the-art while maintaining a reasonable balance between model accuracy and complexity. Code is available at: https://github.com/gudaochangsheng/HSTL.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext dc90f2e9-4fc0-4da2-b97c-bd83d3aa8ae3Cited by top-tier papers10
- Learning Visual Prompt for Gait RecognitionKang Ma, Ying Fu, Chunshui Cao, Saihui Hou et al.CVPR 2024 · 24 citations
- It Takes Two: Accurate Gait Recognition in the Wild via Cross-granularity AlignmentJinkai Zheng, Xinchen Liu, Boyue Zhang, Chenggang Yan et al.ACM MM 2024 · 14 citations
- GLGait: A Global-Local Temporal Receptive Field Network for Gait Recognition in the WildGuozhen Peng, Yunhong Wang, Yuwei Zhao, Shaoxiong Zhang et al.ACM MM 2024 · 12 citations
- Vocabulary-Guided Gait RecognitionPanjian Huang, Saihui Hou, Chunshui Cao, Xu Liu et al.NeurIPS 2025 · 8 citations
- GaitSnippet: Gait Recognition Beyond Unordered Sets and Ordered SequencesSaihui Hou, Chenye Wang, Wenpeng Lang, Zhengxiang Lan et al.ICLR 2026 · 5 citations
Builds on13
- Gait Recognition via Effective Global-Local Feature Representation and Local Temporal AggregationBeibei Lin, Shunli Zhang, Xin YuICCV 2021 · 325 citations
- Gait Recognition in the Wild with Dense 3D Representations and A BenchmarkJinkai Zheng, Xinchen Liu, Wu Liu, Lingxiao He et al.CVPR 2022 · 228 citations
- Gait Recognition with Multiple-Temporal-Scale 3D Convolutional Neural NetworkBeibei Lin, Shunli Zhang, Feng BaoACM MM 2020 · 173 citations
- Context-Sensitive Temporal Feature Learning for Gait RecognitionXiaohu Huang, Duowang Zhu, Hao Wang, Xinggang Wang et al.ICCV 2021 · 159 citations
- Gait Recognition in the Wild: A BenchmarkICCV 2021 · 102 citations
Related papers
- DyGait: Exploiting Dynamic Representations for High-performance Gait RecognitionMing Wang, Xianda Guo, Beibei Lin, Tian Yang et al.ICCV 2023 · 81 citations
- GaitPart: Temporal Part-Based Model for Gait RecognitionChao Fan, Yunjie Peng, Chunshui Cao, Xu Liu et al.CVPR 2020
- Multi-modal Gait Recognition via Effective Spatial-Temporal Feature FusionYufeng Cui, Yimei KangCVPR 2023
- HyperGait: Unleashing the Power of Parsing for Gait Recognition in the Wild via HypergraphJinkai Zheng, Jiaqing Wei, Xinxiang Jin, Yaoqi Sun et al.CVPR 2026
- LandmarkGait: Intrinsic Human Parsing for Gait RecognitionZengbin Wang, Saihui Hou, Man Zhang, Xu Liu et al.ACM MM 2023 · 14 citations
