Optimizing Network Structure for 3D Human Pose Estimation
Hai Ci, Chunyu Wang, Xiaoxuan Ma, Yizhou Wang
Abstract
A human pose is naturally represented as a graph where the joints are the nodes and the bones are the edges. So it is natural to apply Graph Convolutional Network (GCN) to estimate 3D poses from 2D poses. In this work, we propose a generic formulation where both GCN and Fully Connected Network (FCN) are its special cases. From this formulation, we discover that GCN has limited representation power when used for estimating 3D poses. We overcome the limitation by introducing Locally Connected Network (LCN) which is naturally implemented by this generic formulation. It notably improves the representation capability over GCN. In addition, since every joint is only connected to a few joints in its neighborhood, it has strong generalization power. The experiments on public datasets show it: (1) outperforms the state-of-the-arts; (2) is less data hungry than alternative models; (3) generalizes well to unseen actions and datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4d955485-8ecb-463d-b74d-11b52d7e62feCited by top-tier papers59
- 3D Human Pose Estimation with Spatial and Temporal TransformersCe Zheng, Sijie Zhu, Matías Mendieta, Taojiannan Yang et al.ICCV 2021 · 648 citations
- MixSTE: Seq2seq Mixed Spatio-Temporal Encoder for 3D Human Pose Estimation in VideoJinlu Zhang, Zhigang Tu, Jianyu Yang, Yujin Chen et al.CVPR 2022 · 356 citations
- MotionBERT: A Unified Perspective on Learning Human Motion RepresentationsWentao Zhu, Xiaoxuan Ma, Zhaoyang Liu, Libin Liu et al.ICCV 2023 · 322 citations
- Skeleton-aware networks for deep motion retargetingKfir Aberman, Peizhuo Li, Dani Lischinski, Olga Sorkine-Hornung et al.SIGGRAPH 2020 · 210 citations
- Modulated Graph Convolutional Network for 3D Human Pose EstimationZhiming Zou, Wei TangICCV 2021 · 166 citations
Related papers
- GLA-GCN: Global-local Adaptive Graph Convolutional Network for 3D Human Pose Estimation from Monocular VideoBruce X. B. Yu, Zhi Zhang, Yongxu Liu, Sheng-Hua Zhong et al.ICCV 2023 · 131 citations
- Graph Stacked Hourglass Networks for 3D Human Pose EstimationTianhan Xu, Wataru TakanoCVPR 2021
- Conditional Directed Graph Convolution for 3D Human Pose EstimationWenbo Hu, Changgong Zhang, Fangneng Zhan, Lei Zhang et al.ACM MM 2021 · 123 citations
- Graph and Temporal Convolutional Networks for 3D Multi-person Pose Estimation in Monocular VideosYu Cheng, Bo Wang, Bo Yang, Robby T. TanAAAI 2021 · 55 citations
- Learning Dynamic Relationships for 3D Human Motion PredictionQiongjie Cui, Huaijiang Sun, Fei YangCVPR 2020
