View-normalized Skeleton Generation for Action Recognition
Qingzhe Pan, Zhifu Zhao, Xuemei Xie, Jianan Li, Yuhan Cao, Guangming Shi
Abstract
Skeleton-based action recognition has attracted great interest due to low cost of skeleton data acquisition and high robustness to external conditions. A challenging problem of skeleton-based action recognition is the large intra-class gap caused by various viewpoints of skeleton data, which makes the action modeling difficult for network. To alleviate this problem, a feasible solution is to utilize label supervised methods to learn a view-normalization model. However, since the skeleton data in real scenes is acquired from diverse viewpoints, it is difficult to obtain the corresponding view-normalized skeleton as label. Therefore, how to learn a view-normalization model without the supervised label is the key to solving view-variance problem. To this end, we propose a view normalization-based action recognition framework, which is composed of view-normalization generative adversarial network (VN-GAN) and classification network. For VN-GAN, the model is designed to learn the mapping from diverse-view distribution to normalized-view distribution. In detail, it is implemented by graph convolution, where the generator predicts the transformation angles for view normalization and discriminator classifies the real input samples from the generated ones. For classification network, view-normalized data is processed to predict the action class. Without the interference of view variances, classification network can extract more discriminative feature of action. Furthermore, by combining the joint and bone modalities, the proposed method reaches the state-of-the-art performance on NTU RGB+D and NTU-120 RGB+D datasets. Especially in NTU-120 RGB+D, the accuracy is improved by 3.2% and 2.3% under cross-subject and cross-set criteria, respectively.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 56f3400d-bc57-4eb7-9ac3-0eb05712b243Cited by top-tier papers2
- FourLLIE: Boosting Low-Light Image Enhancement by Fourier Frequency InformationChenxi Wang, Hongjun Wu, Zhi JinACM MM 2023 · 214 citations
- Brighten-and-Colorize: A Decoupled Network for Customized Low-Light Image EnhancementChenxi Wang, Zhi JinACM MM 2023 · 26 citations
Related papers
- Gamba: Mamba-based graph convolutional network with dynamic graph topology learning for action recognitionRouyi Zhou, 漾之 吴, Jiajun Wen, Can Gao et al.CVPR 2026
- Generative Action Description Prompts for Skeleton-based Action RecognitionWangmeng Xiang, Chao Li, Yuxuan Zhou, Biao Wang et al.ICCV 2023 · 84 citations
- InfoGCN: Representation Learning for Human Skeleton-based Action RecognitionHyung-Gun Chi, Myoung Hoon Ha, Seung-geun Chi, Sang Wan Lee et al.CVPR 2022 · 383 citations
- Dynamic Semantic-Based Spatial Graph Convolution Network for Skeleton-Based Human Action RecognitionJianyang Xie, Yanda Meng, Yitian Zhao, Anh Nguyen et al.AAAI 2024 · 59 citations
- Topology-Aware Convolutional Neural Network for Efficient Skeleton-Based Action RecognitionKailin Xu, Fanfan Ye, Qiaoyong Zhong, Di XieAAAI 2022 · 168 citations
