Cross View Fusion for 3D Human Pose Estimation
Haibo Qiu, Chunyu Wang, Jingdong Wang, Naiyan Wang, Wenjun Zeng
Abstract
We present an approach to recover absolute 3D human poses from multi-view images by incorporating multi-view geometric priors in our model. It consists of two separate steps: (1) estimating the 2D poses in multi-view images and (2) recovering the 3D poses from the multi-view 2D poses. First, we introduce a cross-view fusion scheme into CNN to jointly estimate 2D poses for multiple views. Consequently, the 2D pose estimation for each view already benefits from other views. Second, we present a recursive Pictorial Structure Model to recover the 3D pose from the multi-view 2D poses. It gradually improves the accuracy of 3D pose with affordable computational cost. We test our method on two public datasets H36M and Total Capture. The Mean Per Joint Position Errors on the two datasets are 26mm and 29mm, which outperforms the state-of-the-arts remarkably (26mm vs 52mm, 29mm vs 35mm). Our code is released at https:// github.com/ microsoft/ multiview-human-pose-estimation-pytorch.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 481ff84f-e8fe-40b5-8ae2-1b9033d406acCited by top-tier papers45
- AI Choreographer: Music Conditioned 3D Dance Generation with AIST++Ruilong Li, Shan Yang, David A. Ross, Angjoo KanazawaICCV 2021 · 701 citations
- SPEC: Seeing People in the Wild with an Estimated CameraMuhammed Kocabas, Chun-Hao P. Huang, Joachim Tesch, Lea Müller et al.ICCV 2021 · 181 citations
- Direct Multi-view Multi-person 3D Pose EstimationTao Wang, Jianfeng Zhang, Yujun Cai, Shuicheng Yan et al.NeurIPS 2021 · 147 citations
- Conditional Directed Graph Convolution for 3D Human Pose EstimationWenbo Hu, Changgong Zhang, Fangneng Zhan, Lei Zhang et al.ACM MM 2021 · 123 citations
- GLAMR: Global Occlusion-Aware Human Mesh Recovery with Dynamic CamerasYe Yuan, Umar Iqbal, Pavlo Molchanov, Kris Kitani et al.CVPR 2022 · 111 citations
Related papers
- Learnable Triangulation of Human PoseKarim Iskakov, Egor Burkov, Victor S. Lempitsky, Yury MalkovICCV 2019 · 419 citations
- Fusing Wearable IMUs With Multi-View Images for Human Pose Estimation: A Geometric ApproachZhe Zhang, Chunyu Wang, Wenhu Qin, Wenjun ZengCVPR 2020
- Cross-View Tracking for Multi-Human 3D Pose Estimation at Over 100 FPSLong Chen, Haizhou Ai, Rui Chen, Zijie Zhuang et al.CVPR 2020
- Mocap-2-to-3: Multi-view Lifting for Monocular Motion Recovery with 2D PretrainingZhumei Wang, Zechen Hu, Ruoxi Guo, Huaijin Pi et al.CVPR 2026 · 1 citation
- Towards Stable Human Pose Estimation via Cross-View Fusion and Foot StabilizationLi'an Zhuo, Jian Cao, Qi Wang, Bang Zhang et al.CVPR 2023
