Learning Multi-Granular Spatio-Temporal Graph Network for Skeleton-based Action Recognition
Tailin Chen, Desen Zhou, Jian Wang, Shidong Wang, Yu Guan, Xuming He, Errui Ding
Abstract
The task of skeleton-based action recognition remains a core challenge in human-centred scene understanding due to the multiple granularities and large variation in human motion. Existing approaches typically employ a single neural representation for different motion patterns, which has difficulty in capturing fine-grained action classes given limited training data. To address the aforementioned problems, we propose a novel multi-granular spatio-temporal graph network for skeleton-based action classification that jointly models the coarse- and fine-grained skeleton motion patterns. To this end, we develop a dual-head graph network consisting of two interleaved branches, which enables us to extract features at two spatio-temporal resolutions in an effective and efficient manner. Moreover, our network utilises a cross-head communication strategy to mutually enhance the representations of both heads. We conducted extensive experiments on three large-scale datasets, namely NTU RGB+D 60, NTU RGB+D 120, and Kinetics-Skeleton, and achieves the state-of-the-art performance on all the benchmarks, which validates the effectiveness of our method1.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 05a51f58-9ff9-4241-9cf9-5f476400ac1fCited by top-tier papers8
- SkeletonMAE: Graph-based Masked Autoencoder for Skeleton Sequence Pre-trainingHong Yan, Yang Liu, Yushen Wei, Zhen Li et al.ICCV 2023 · 77 citations
- Cross-Modal Learning with 3D Deformable Attention for Action RecognitionSangwon Kim, Dasom Ahn, ByoungChul KoICCV 2023 · 49 citations
- Multi-Modality Co-Learning for Efficient Skeleton-based Action RecognitionJinfu Liu, Chen Chen, Mengyuan LiuACM MM 2024 · 27 citations
- Shifting Perspective to See Difference: A Novel Multi-view Method for Skeleton based Action RecognitionRuijie Hou, Yanran Li, Ningyu Zhang, Yulin Zhou et al.ACM MM 2022 · 18 citations
- Behavioral Recognition of Skeletal Data Based on Targeted Dual Fusion StrategyXiao Yun, Chenglong Xu, Kévin Riou, Kaiwen Dong et al.AAAI 2024 · 14 citations
Builds on12
- SlowFast Networks for Video RecognitionChristoph Feichtenhofer, Haoqi Fan, Jitendra Malik, Kaiming HeICCV 2019 · 4,104 citations
- Learning Graph Convolutional Network for Skeleton-Based Human Action Recognition by Neural SearchingWei Peng, Xiaopeng Hong, Haoyu Chen, Guoying ZhaoAAAI 2020 · 362 citations
- Stronger, Faster and More Explainable: A Graph Convolutional Baseline for Skeleton-based Action RecognitionYi-Fan Song, Zhang Zhang, Caifeng Shan, Liang WangACM MM 2020 · 361 citations
- Dynamic GCN: Context-enriched Topology Learning for Skeleton-based Action RecognitionFanfan Ye, Shiliang Pu, Qiaoyong Zhong, Chao Li et al.ACM MM 2020 · 348 citations
- Multi-Scale Spatial Temporal Graph Convolutional Network for Skeleton-Based Action RecognitionZhan Chen, Sicheng Li, Bing Yang, Qinghan Li et al.AAAI 2021 · 341 citations
Related papers
- Dynamic Semantic-Based Spatial Graph Convolution Network for Skeleton-Based Human Action RecognitionJianyang Xie, Yanda Meng, Yitian Zhao, Anh Nguyen et al.AAAI 2024 · 59 citations
- Revealing Key Details to See Differences: A Novel Prototypical Perspective for Skeleton-based Action RecognitionHongda Liu, Yunfan Liu, Min Ren, Hao Wang et al.CVPR 2025
- Skeleton-based Human Action Recognition via Large-kernel Attention Graph Convolutional NetworkYanan Liu, Hao Zhang, Yanqiu Li, Kangjian He et al.IEEE VR 2023 · 123 citations
- Disentangling and Unifying Graph Convolutions for Skeleton-Based Action RecognitionZiyu Liu, Hongwen Zhang, Zhenghao Chen, Zhiyong Wang et al.CVPR 2020
- Kinematic Enhanced Hypergraph Convolutional Network for Skeleton-based Human Action Recognition with LLM Training GuidesNan Ma, Beining Sun, Yiheng Han, Genbao XuACM MM 2025 · 3 citations
