View-normalized Skeleton Generation for Action Recognition
Qingzhe Pan, Zhifu Zhao, Xuemei Xie, Jianan Li, Yuhan Cao, Guangming Shi
摘要
Skeleton-based action recognition has attracted great interest due to low cost of skeleton data acquisition and high robustness to external conditions. A challenging problem of skeleton-based action recognition is the large intra-class gap caused by various viewpoints of skeleton data, which makes the action modeling difficult for network. To alleviate this problem, a feasible solution is to utilize label supervised methods to learn a view-normalization model. However, since the skeleton data in real scenes is acquired from diverse viewpoints, it is difficult to obtain the corresponding view-normalized skeleton as label. Therefore, how to learn a view-normalization model without the supervised label is the key to solving view-variance problem. To this end, we propose a view normalization-based action recognition framework, which is composed of view-normalization generative adversarial network (VN-GAN) and classification network. For VN-GAN, the model is designed to learn the mapping from diverse-view distribution to normalized-view distribution. In detail, it is implemented by graph convolution, where the generator predicts the transformation angles for view normalization and discriminator classifies the real input samples from the generated ones. For classification network, view-normalized data is processed to predict the action class. Without the interference of view variances, classification network can extract more discriminative feature of action. Furthermore, by combining the joint and bone modalities, the proposed method reaches the state-of-the-art performance on NTU RGB+D and NTU-120 RGB+D datasets. Especially in NTU-120 RGB+D, the accuracy is improved by 3.2% and 2.3% under cross-subject and cross-set criteria, respectively.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper2
- FourLLIE: Boosting Low-Light Image Enhancement by Fourier Frequency InformationChenxi Wang, Hongjun Wu, Zhi JinACM MM 2023 · 被引用 214 次
- Brighten-and-Colorize: A Decoupled Network for Customized Low-Light Image EnhancementChenxi Wang, Zhi JinACM MM 2023 · 被引用 26 次
相关 Paper
- Gamba: Mamba-based graph convolutional network with dynamic graph topology learning for action recognitionRouyi Zhou, 漾之 吴, Jiajun Wen, Can Gao 等CVPR 2026
- Generative Action Description Prompts for Skeleton-based Action RecognitionWangmeng Xiang, Chao Li, Yuxuan Zhou, Biao Wang 等ICCV 2023 · 被引用 84 次
- InfoGCN: Representation Learning for Human Skeleton-based Action RecognitionHyung-Gun Chi, Myoung Hoon Ha, Seung-geun Chi, Sang Wan Lee 等CVPR 2022 · 被引用 383 次
- Dynamic Semantic-Based Spatial Graph Convolution Network for Skeleton-Based Human Action RecognitionJianyang Xie, Yanda Meng, Yitian Zhao, Anh Nguyen 等AAAI 2024 · 被引用 59 次
- Topology-Aware Convolutional Neural Network for Efficient Skeleton-Based Action RecognitionKailin Xu, Fanfan Ye, Qiaoyong Zhong, Di XieAAAI 2022 · 被引用 168 次
