Learning Visual Prompt for Gait Recognition
Kang Ma, Ying Fu, Chunshui Cao, Saihui Hou, Yongzhen Huang, Dezhi Zheng
Abstract
Gait, a prevalent and complex form of human motion, plays a significant role in the field of long-range pedestrian retrieval due to the unique characteristics inherent in individual motion patterns. However, gait recognition in real-world scenarios is challenging due to the limitations of capturing comprehensive cross-viewing and crossclothing data. Additionally, distractors such as occlusions, directional changes, and lingering movements further complicate the problem. The widespread application of deep learning techniques has led to the development of various potential gait recognition methods. However, these methods utilize convolutional networks to extract shared information across different views and attire conditions. Once trained, the parameters and non-linear function become constrained to fixed patterns, limiting their adaptability to various distractors in real-world scenarios. In this paper, we present a unified gait recognition framework to extract global motion patterns and develop a novel dynamic transformer to generate representative gait features. Specifically, we develop a trainable part-based prompt pool with numerous key-value pairs that can dynamically select prompt templates to incorporate into the gait sequence, thereby providing task-relevant shared knowledge information. Furthermore, we specifically design dynamic attention to extract robust motion patterns and address the length generalization issue. Extensive experiments on four widely recognized gait datasets, i.e., Gait3D, GREW, OUMVLP, and CASIA-B, reveal that the proposed method yields substantial improvements compared to current state-of-the-art approaches.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e50eee09-1d23-4e5e-82d7-5a458858667dCited by top-tier papers10
- Vocabulary-Guided Gait RecognitionPanjian Huang, Saihui Hou, Chunshui Cao, Xu Liu et al.NeurIPS 2025 · 8 citations
- GaitSnippet: Gait Recognition Beyond Unordered Sets and Ordered SequencesSaihui Hou, Chenye Wang, Wenpeng Lang, Zhengxiang Lan et al.ICLR 2026 · 5 citations
- Exploring Task-Level Optimal Prompts for Visual In-Context LearningYan Zhu, Huan Ma, Changqing ZhangAAAI 2025 · 4 citations
- Learning a Unified Template for Gait RecognitionPanjian Huang, Saihui Hou, Junzhou Huang, Yongzhen HuangICCV 2025 · 2 citations
- MMGait: Towards Multi-Modal Gait RecognitionChenye Wang, Qingyuan Cai, Saihui Hou, Aoqi Li et al.CVPR 2026 · 1 citation
Builds on24
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- TSM: Temporal Shift Module for Efficient Video UnderstandingJi Lin, Chuang Gan, Song HanICCV 2019 · 2,049 citations
- Efficient Streaming Language Models with Attention SinksGuangxuan Xiao, Yuandong Tian, Beidi Chen, Song Han et al.ICLR 2024 · 1,714 citations
Related papers
- GLGait: A Global-Local Temporal Receptive Field Network for Gait Recognition in the WildGuozhen Peng, Yunhong Wang, Yuwei Zhao, Shaoxiong Zhang et al.ACM MM 2024 · 12 citations
- Dynamic Aggregated Network for Gait RecognitionKang Ma, Ying Fu, Dezhi Zheng, Chunshui Cao et al.CVPR 2023
- GPGait: Generalized Pose-based Gait RecognitionYang Fu, Shibei Meng, Saihui Hou, Xuecai Hu et al.ICCV 2023 · 86 citations
- DyGait: Exploiting Dynamic Representations for High-performance Gait RecognitionMing Wang, Xianda Guo, Beibei Lin, Tian Yang et al.ICCV 2023 · 81 citations
- Hierarchical Spatio-Temporal Representation Learning for Gait RecognitionLei Wang, Bo Liu, Fangfang Liang, Bincheng WangICCV 2023 · 43 citations
