Mix Dimension in Poincaré Geometry for 3D Skeleton-based Action Recognition
Wei Peng, Jingang Shi, Zhaoqiang Xia, Guoying Zhao
摘要
Graph Convolutional Networks (GCNs) have already demonstrated their powerful ability to model the irregular data, e.g., skeletal data in human action recognition, providing an exciting new way to fuse rich structural information for nodes residing in different parts of a graph. In human action recognition, current works introduce a dynamic graph generation mechanism to better capture the underlying semantic skeleton connections and thus improves the performance. In this paper, we provide an orthogonal way to explore the underlying connections. Instead of introducing an expensive dynamic graph generation paradigm, we build a more efficient GCN on a Riemann manifold, which we think is a more suitable space to model the graph data, to make the extracted representations fit the embedding matrix. Specifically, we present a novel spatial-temporal GCN (ST-GCN) architecture which is defined via the Poincaré geometry such that it is able to better model the latent anatomy of the structure data. To further explore the optimal projection dimension in the Riemann space, we mix different dimensions on the manifold and provide an efficient way to explore the dimension for each ST-GCN layer. With the final resulted architecture, we evaluate our method on two current largest scale 3D datasets, i.e., NTU RGB+D and NTU RGB+D 120. The comparison results show that the model could achieve a superior performance under any given evaluation metrics with only 40% model size when compared with the previous best GCN method, which proves the effectiveness of our model.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- Skeleton-Contrastive 3D Action Representation LearningFida Mohammad Thoker, Hazel Doughty, Cees G. M. SnoekACM MM 2021 · 被引用 158 次
- Learning Skeletal Graph Neural Networks for Hard 3D Pose EstimationAiling Zeng, Xiao Sun, Lei Yang, Nanxuan Zhao 等ICCV 2021 · 被引用 146 次
- Skeleton-based Action Recognition via Adaptive Cross-Form LearningXuanhan Wang, Yan Dai, Lianli Gao, Jingkuan SongACM MM 2022 · 被引用 28 次
- Beyond Euclidean: Dual-Space Representation Learning for Weakly Supervised Video Violence DetectionJiaxu Leng, Zhanjie Wu, Mingpi Tan, Yiran Liu 等NeurIPS 2024 · 被引用 22 次
- κHGCN: Tree-likeness Modeling via Continuous and Discrete Curvature LearningMenglin Yang, Min Zhou, Lujia Pan, Irwin KingKDD 2023 · 被引用 14 次
它引用的顶会 Paper2
相关 Paper
- Dynamic GCN: Context-enriched Topology Learning for Skeleton-based Action RecognitionFanfan Ye, Shiliang Pu, Qiaoyong Zhong, Chao Li 等ACM MM 2020 · 被引用 348 次
- Multi-Scale Spatial Temporal Graph Convolutional Network for Skeleton-Based Action RecognitionZhan Chen, Sicheng Li, Bing Yang, Qinghan Li 等AAAI 2021 · 被引用 341 次
- Disentangling and Unifying Graph Convolutions for Skeleton-Based Action RecognitionZiyu Liu, Hongwen Zhang, Zhenghao Chen, Zhiyong Wang 等CVPR 2020
- Leveraging Spatio-Temporal Dependency for Skeleton-Based Action RecognitionJungho Lee, Minhyeok Lee, Suhwan Cho, Sungmin Woo 等ICCV 2023 · 被引用 28 次
- Hierarchically Decomposed Graph Convolutional Networks for Skeleton-Based Action RecognitionJungho Lee, Minhyeok Lee, Dogyoon Lee, Sangyoun LeeICCV 2023 · 被引用 236 次
