Unlocking Motion from Large Vision Models with a Semantic and Kinematic Duality for Gait Recognition
Zhanbo Huang, Dingqiang Ye, Xiaoming Liu, Yu Kong
Abstract
Existing set-based gait recognition methods achieve remarkable performance by capturing global semantic context. However, their order-invariant nature prevents them from modeling the fine-grained kinematic patterns that unfold over time. To unify the global and process-level representations, we propose GaitMax, a framework capturing both semantic context and kinematic motion. GaitMax leverages attention-based spatiotemporal modeling to dynamically represent detailed part-level trajectories. While this detailed representation is more powerful, it also captures more nuisance factors (e.g., clothing, viewpoint), leading to potential shortcuts. To mitigate this, we introduce Conditional Decorrelation Loss (CDLoss), which explicitly disentangles the gait embeddings from nuisance factors using vision-language supervision. This loss requires high-quality nuisance descriptions. We therefore construct GCaption, a new resource that provides natural language annotations for multiple gait datasets, moving beyond simple categorical labels. GCaption not only enables our CDLoss but also serves as a foundation for future context-aware gait analysis. Models, code, and resources are available at https://zbhuang.com/gait-max.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b2b962d0-d50f-42b4-b4ec-83d050e874e4Cited by top-tier papers1
Ask how each one uses itBuilds on20
- Gait Recognition via Effective Global-Local Feature Representation and Local Temporal AggregationBeibei Lin, Shunli Zhang, Xin YuICCV 2021 · 325 citations
- Gait Recognition in the Wild with Dense 3D Representations and A BenchmarkJinkai Zheng, Xinchen Liu, Wu Liu, Lingxiao He et al.CVPR 2022 · 228 citations
- Gait Recognition in the Wild: A BenchmarkICCV 2021 · 102 citations
- SkeletonGait: Gait Recognition Using Skeleton MapsChao Fan, Jingzhe Ma, Dongyang Jin, Chuanfu Shen et al.AAAI 2024 · 92 citations
- BigGait: Learning Gait Representation You Want by Large Vision ModelsDingqiang Ye, Chao Fan, Jingzhe Ma, Xiaoming Liu et al.CVPR 2024 · 40 citations
Related papers
- Language-Guided and Motion-Aware Gait Representation for Generalizable RecognitionZhengxian Wu, Chuanrui Zhang, Shenao Jiang, Hangrui Xu et al.AAAI 2026 · 1 citation
- Gait Recognition via Semi-supervised Disentangled Representation Learning to Identity and Covariate FeaturesXiang Li, Yasushi Makihara, Chi Xu, Yasushi Yagi et al.CVPR 2020
- HybridGait: A Benchmark for Spatial-Temporal Cloth-Changing Gait Recognition with Hybrid ExplorationsYilan Dong, Chunlin Yu, Ruiyang Ha, Ye Shi et al.AAAI 2024 · 31 citations
- Learning Visual Prompt for Gait RecognitionKang Ma, Ying Fu, Chunshui Cao, Saihui Hou et al.CVPR 2024 · 24 citations
- On Denoising Walking Videos for Gait RecognitionDongyang Jin, Chao Fan, Jingzhe Ma, Jingkai Zhou et al.CVPR 2025
