Decoupled Representation Learning for Skeleton-Based Gesture Recognition
Jianbo Liu, Yongcheng Liu, Ying Wang, Véronique Prinet, Shiming Xiang, Chunhong Pan
摘要
Skeleton-based gesture recognition is very challenging, as the high-level information in gesture is expressed by a sequence of complexly composite motions. Previous works often learn all the motions with a single model. In this paper, we propose to decouple the gesture into hand posture variations and hand movements, which are then modeled separately. For the former, the skeleton sequence is embedded into a 3D hand posture evolution volume (HPEV) to represent fine-grained posture variations. For the latter, the shifts of hand center and fingertips are arranged as a 2D hand movement map (HMM) to capture holistic movements. To learn from the two inhomogeneous representations for gesture recognition, we propose an end-to-end two-stream network. The HPEV stream integrates both spatial layout and temporal evolution information of hand postures by a dedicated 3D CNN, while the HMM stream develops an efficient 2D CNN to extract hand movement features. Eventually, the predictions of the two streams are aggregated with high efficiency. Extensive experiments on SHREC'17 Track, DHG-14/28 and FPHA datasets demonstrate that our method is competitive with the state-of-the-art.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它相关 Paper
- An Efficient PointLSTM for Point Clouds Based Gesture RecognitionYuecong Min, Yanxiao Zhang, Xiujuan Chai, Xilin ChenCVPR 2020
- Back-Hand-Pose: 3D Hand Pose Estimation for a Wrist-worn Camera via Dorsum Deformation NetworkErwin Wu, Ye Yuan, Hui-Shyong Yeo, Aaron Quigley 等UIST 2020 · 被引用 73 次
- Diverse 3D Hand Gesture Prediction from Body Dynamics by Bilateral Hand DisentanglementXingqun Qi, Chen Liu, Muyi Sun, Lincheng Li 等CVPR 2023
- HandFoldingNet: A 3D Hand Pose Estimation Network Using Multiscale-Feature Guided Folding of a 2D Hand SkeletonWencan Cheng, Jae Hyun Park, Jong Hwan KoICCV 2021 · 被引用 50 次
- HandVoxNet: Deep Voxel-Based Network for 3D Hand Shape and Pose Estimation From a Single Depth MapJameel Malik, Ibrahim Abdelaziz, Ahmed Elhayek, Soshi Shimada 等CVPR 2020
