BiggerGait: Unlocking Gait Recognition with Layer-wise Representations from Large Vision Models
Dingqiang Ye, Chao Fan, Zhanbo Huang, Chengwen Luo, Jianqiang Li, Shiqi Yu, Xiaoming Liu
Abstract
Large vision models (LVM) based gait recognition has achieved impressive performance. However, existing LVM-based approaches may overemphasize gait priors while neglecting the intrinsic value of LVM itself, particularly the rich, distinct representations across its multi-layers. To adequately unlock LVM's potential, this work investigates the impact of layer-wise representations on downstream recognition tasks. Our analysis reveals that LVM's intermediate layers offer complementary properties across tasks, integrating them yields an impressive improvement even without rich well-designed gait priors. Building on this insight, we propose a simple and universal baseline for LVM-based gait recognition, termed BiggerGait. Comprehensive evaluations on CCPG, CAISA-B*, SUSTech1K, and CCGR_MINI validate the superiority of BiggerGait across both within-and cross-domain tasks, establishing it as a simple yet practical baseline for gait representation learning. All the models and code are available at https://github.com/ShiqiYu/OpenGait/.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 57a115b2-d3d3-4493-93fc-a088746126b6Cited by top-tier papers7
- SCALAR: Scale-wise Controllable Visual Autoregressive LearningRyan Xu, Dongyang Jin, Yancheng Bai, Rui Lan et al.AAAI 2026 · 16 citations
- DINOv2 Driven Gait Representation Learning for Video-Based Visible-Infrared Person Re-identificationYujie Yang, Shuang Li, Jun Ye, Neng Dong et al.ACM MM 2025 · 10 citations
- FusionAgent: A Multimodal Agent with Dynamic Model Selection for Human RecognitionJie Zhu, Xiao Guo, Yiyang Su, Anil K. Jain et al.CVPR 2026 · 7 citations
- Unlocking Motion from Large Vision Models with a Semantic and Kinematic Duality for Gait RecognitionZhanbo Huang, Dingqiang Ye, Xiaoming Liu, Yu KongCVPR 2026 · 4 citations
- EventGait: Towards Robust Gait Recognition with Event StreamsSenyan Xu, Shuai Chen, Chuanfu Shen, Kean Liu et al.CVPR 2026 · 2 citations
Builds on30
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Perception Encoder: The best visual embeddings are not at the output of the networkDaniel Bolya, Po-Yao Huang, Peize Sun, Jang Hyun Cho et al.NeurIPS 2025 · 359 citations
- Gait Recognition via Effective Global-Local Feature Representation and Local Temporal AggregationBeibei Lin, Shunli Zhang, Xin YuICCV 2021 · 325 citations
Related papers
- BigGait: Learning Gait Representation You Want by Large Vision ModelsDingqiang Ye, Chao Fan, Jingzhe Ma, Xiaoming Liu et al.CVPR 2024 · 40 citations
- OpenGait: Revisiting Gait Recognition Toward Better PracticalityChao Fan, Junhao Liang, Chuanfu Shen, Saihui Hou et al.CVPR 2023
- CarGait: Cross-Attention Based Re-ranking for Gait RecognitionGavriel Habib, Noa Barzilay, Or Shimshi, Rami Ben-Ari et al.ICCV 2025
- Hierarchical Spatio-Temporal Representation Learning for Gait RecognitionLei Wang, Bo Liu, Fangfang Liang, Bincheng WangICCV 2023 · 43 citations
- Bridging Gait Recognition and Large Language Models Sequence ModelingShaopeng Yang, Jilong Wang, Saihui Hou, Xu Liu et al.CVPR 2025
