Multi-modal Gait Recognition via Effective Spatial-Temporal Feature Fusion
Yufeng Cui, Yimei Kang
摘要
Gait recognition is a biometric technology that identifies people by their walking patterns. The silhouettesbased method and the skeletons-based method are the two most popular approaches. However, the silhouette data are easily affected by clothing occlusion, and the skeleton data lack body shape information. To obtain a more robust and comprehensive gait representation for recognition, we propose a transformer-based gait recognition framework called MMGaitFormer, which effectively fuses and aggregates the spatial-temporal information from the skeletons and silhouettes. Specifically, a Spatial Fusion Module (SFM) and a Temporal Fusion Module (TFM) are proposed for effective spatial-level and temporal-level feature fusion, respectively. The SFM performs fine-grained body parts spatial fusion and guides the alignment of each part of the silhouette and each joint of the skeleton through the attention mechanism. The TFM performs temporal modeling through Cycle Position Embedding (CPE) and fuses temporal information of two modalities. Experiments demonstrate that our MMGaitFormer achieves state-of-the-art performance on popular gait datasets. For the most challenging "CL" (i.e., walking in different clothes) condition in CASIA-B, our method achieves a rank-1 accuracy of 94.8%, which outperforms the state-of-the-art single-modal methods by a large margin.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- It Takes Two: Accurate Gait Recognition in the Wild via Cross-granularity AlignmentJinkai Zheng, Xinchen Liu, Boyue Zhang, Chenggang Yan 等ACM MM 2024 · 被引用 14 次
- Vocabulary-Guided Gait RecognitionPanjian Huang, Saihui Hou, Chunshui Cao, Xu Liu 等NeurIPS 2025 · 被引用 8 次
- GaitSnippet: Gait Recognition Beyond Unordered Sets and Ordered SequencesSaihui Hou, Chenye Wang, Wenpeng Lang, Zhengxiang Lan 等ICLR 2026 · 被引用 5 次
- WaveLoss: An Adaptive Dynamic Loss for Deep Gait RecognitionZicheng Wang, Qiuxia WuAAAI 2025 · 被引用 4 次
- DepthGait: Multi-Scale Cross-Level Feature Fusion of RGB-Derived Depth and Silhouette Sequences for Robust Gait RecognitionXinzhu Li, Juepeng Zheng, Yikun Chen, Xudong Mao 等ACM MM 2025 · 被引用 2 次
它引用的顶会 Paper4
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Gait Recognition via Effective Global-Local Feature Representation and Local Temporal AggregationBeibei Lin, Shunli Zhang, Xin YuICCV 2021 · 被引用 325 次
- Gait Recognition with Multiple-Temporal-Scale 3D Convolutional Neural NetworkBeibei Lin, Shunli Zhang, Feng BaoACM MM 2020 · 被引用 173 次
- GaitPart: Temporal Part-Based Model for Gait RecognitionChao Fan, Yunjie Peng, Chunshui Cao, Xu Liu 等CVPR 2020
相关 Paper
- GaitCycFormer: Leveraging Gait Cycles and Transformers for Gait Emotion RecognitionQingyang Zeng, Lin ShangAAAI 2025 · 被引用 6 次
- HybridGait: A Benchmark for Spatial-Temporal Cloth-Changing Gait Recognition with Hybrid ExplorationsYilan Dong, Chunlin Yu, Ruiyang Ha, Ye Shi 等AAAI 2024 · 被引用 31 次
- Gait Transformer: End-to-End Transformer Backbone for Gait RecognitionSaihui Hou, Wenpeng Lang, Jilong Wang, Yan Huang 等AAAI 2026
- DyGait: Exploiting Dynamic Representations for High-performance Gait RecognitionMing Wang, Xianda Guo, Beibei Lin, Tian Yang 等ICCV 2023 · 被引用 81 次
- Hierarchical Spatio-Temporal Representation Learning for Gait RecognitionLei Wang, Bo Liu, Fangfang Liang, Bincheng WangICCV 2023 · 被引用 43 次
