Modulated Graph Convolutional Network for 3D Human Pose Estimation
Zhiming Zou, Wei Tang
Abstract
The graph convolutional network (GCN) has recently achieved promising performance of 3D human pose estimation (HPE) by modeling the relationship among body parts. However, most prior GCN approaches suffer from two main drawbacks. First, they share a feature transformation for each node within a graph convolution layer. This prevents them from learning different relations between different body joints. Second, the graph is usually defined according to the human skeleton and is suboptimal because human activities often exhibit motion patterns beyond the natural connections of body joints. To address these limitations, we introduce a novel Modulated GCN for 3D HPE. It consists of two main components: weight modulation and affinity modulation. Weight modulation learns different modulation vectors for different nodes so that the feature transformations of different nodes are disentangled while retaining a small model size. Affinity modulation adjusts the graph structure in a GCN so that it can model additional edges beyond the human skeleton. We investigate several affinity modulation methods as well as the impact of regularizations. Rigorous ablation study indicates both types of modulation improve performance with negligible overhead. Compared with state-of-the-art GCNs for 3D HPE, our approach either significantly reduces the estimation errors, e.g., by around 10%, while retaining a small model size or drastically reduces the model size, e.g., from 4.22M to 0.29M (a 14.5× reduction), while achieving comparable performance. Results on two benchmarks show our Modulated GCN outperforms some recent states of the art. Our code is available at https://github.com/ ZhimingZo/Modulated-GCN .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers21
- MHFormer: Multi-Hypothesis Transformer for 3D Human Pose EstimationWenhao Li, Hong Liu, Hao Tang, Pichao Wang et al.CVPR 2022 · 403 citations
- Pose-Oriented Transformer with Uncertainty-Guided Refinement for 2D-to-3D Human Pose EstimationHan Li, Bowen Shi, Wenrui Dai, Hongwei Zheng et al.AAAI 2023 · 76 citations
- KTPFormer: Kinematics and Trajectory Prior Knowledge-Enhanced Transformer for 3D Human Pose EstimationJihua Peng, Yanghong Zhou, P. Y. MokCVPR 2024 · 67 citations
- CEE-Net: Complementary End-to-End Network for 3D Human Pose Generation and EstimationHaolun Li, Chi-Man PunAAAI 2023 · 46 citations
- FinePOSE: Fine-Grained Prompt-Driven 3D Human Pose Estimation via Diffusion ModelsJinglin Xu, Yijie Guo, Yuxin PengCVPR 2024 · 39 citations
Builds on9
- Exploiting Spatial-Temporal Relationships for 3D Pose Estimation via Graph Convolutional NetworksYujun Cai, Liuhao Ge, Jun Liu, Jianfei Cai et al.ICCV 2019 · 504 citations
- Relation-Aware Graph Attention Network for Visual Question AnsweringLinjie Li, Zhe Gan, Yu Cheng, Jingjing LiuICCV 2019 · 391 citations
- Optimizing Network Structure for 3D Human Pose EstimationHai Ci, Chunyu Wang, Xiaoxuan Ma, Yizhou WangICCV 2019 · 267 citations
- Monocular 3D Human Pose Estimation by Generation and Ordinal RankingSaurabh Sharma, Pavan Teja Varigonda, Prashast Bindal, Abhishek Sharma et al.ICCV 2019 · 177 citations
- Bayesian Graph Convolution LSTM for Skeleton Based Action RecognitionRui Zhao, Kang Wang, Hui Su, Qiang JiICCV 2019 · 104 citations
Related papers
- Graph Stacked Hourglass Networks for 3D Human Pose EstimationTianhan Xu, Wataru TakanoCVPR 2021
- Part-Level Graph Convolutional Network for Skeleton-Based Action RecognitionLinjiang Huang, Yan Huang, Wanli Ouyang, Liang WangAAAI 2020 · 111 citations
- HopFIR: Hop-wise GraphFormer with Intragroup Joint Refinement for 3D Human Pose EstimationKai Zhai, Qiang Nie, Bo Ouyang, Xiang Li et al.ICCV 2023 · 15 citations
- Revisiting Skeleton-based Action RecognitionHaodong Duan, Yue Zhao, Kai Chen, Dahua Lin et al.CVPR 2022 · 752 citations
- Space-Time-Separable Graph Convolutional Network for Pose ForecastingTheodoros Sofianos, Alessio Sampieri, Luca Franco, Fabio GalassoICCV 2021 · 188 citations
