Monocular 3D Multi-Person Pose Estimation by Integrating Top-Down and Bottom-Up Networks
Yu Cheng, Bo Wang, Bo Yang, Robby T. Tan
摘要
In monocular video 3D multi-person pose estimation, inter-person occlusion and close interactions can cause human detection to be erroneous and human-joints grouping to be unreliable. Existing top-down methods rely on human detection and thus suffer from these problems. Existing bottom-up methods do not use human detection, but they process all persons at once at the same scale, causing them to be sensitive to multiple-persons scale variations. To address these challenges, we propose the integration of top-down and bottom-up approaches to exploit their strengths. Our top-down network estimates human joints from all persons instead of one in an image patch, making it robust to possible erroneous bounding boxes. Our bottomup network incorporates human-detection based normal- ized heatmaps, allowing the network to be more robust in handling scale variations. Finally, the estimated 3D poses from the top-down and bottom-up networks are fed into our integration network for final 3D poses. Besides the integration of top-down and bottom-up networks, unlike existing pose discriminators that are designed solely for a single person, and consequently cannot assess natural interperson interactions, we propose a two-person pose discriminator that enforces natural two-person interactions. Lastly, we also apply a semi-supervised method to overcome the 3D ground-truth data scarcity. Quantitative and qualitative evaluations show the effectiveness of the proposed method. Our code is available publicly. 1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- Single-Stage is Enough: Multi-Person Absolute 3D Pose EstimationLei Jin, Chenyang Xu, Xiaojuan Wang, Yabo Xiao 等CVPR 2022 · 被引用 40 次
- Reconstructing Groups of People with Hypergraph Relational ReasoningBuzhen Huang, Jingyi Ju, Zhihao Li, Yangang WangICCV 2023 · 被引用 21 次
- Towards Robust and Smooth 3D Multi-Person Pose Estimation from Monocular Videos in the WildSungchan Park, Eunyi You, Inhoe Lee, Joonseok LeeICCV 2023 · 被引用 16 次
- Mutual Adaptive Reasoning for Monocular 3D Multi-Person Pose EstimationJuze Zhang, Jingya Wang, Ye Shi, Fei Gao 等ACM MM 2022 · 被引用 15 次
- PhysPT: Physics-aware Pretrained Transformer for Estimating Human Dynamics from Monocular VideosYufei Zhang, Jeffrey O. Kephart, Zijun Cui, Qiang JiCVPR 2024 · 被引用 14 次
它引用的顶会 Paper13
- Learning to Reconstruct 3D Human Pose and Shape via Model-Fitting in the LoopNikos Kolotouros, Georgios Pavlakos, Michael J. Black, Kostas DaniilidisICCV 2019 · 被引用 1,139 次
- Exploiting Spatial-Temporal Relationships for 3D Pose Estimation via Graph Convolutional NetworksYujun Cai, Liuhao Ge, Jun Liu, Jianfei Cai 等ICCV 2019 · 被引用 504 次
- Camera Distance-Aware Top-Down Approach for 3D Multi-Person Pose Estimation From a Single RGB ImageGyeongsik Moon, Ju Yong Chang, Kyoung Mu LeeICCV 2019 · 被引用 368 次
- XNect: real-time multi-person 3D motion capture with a single RGB cameraDushyant Mehta, Oleksandr Sotnychenko, Franziska Mueller, Weipeng Xu 等SIGGRAPH 2020 · 被引用 267 次
- Occlusion-Aware Networks for 3D Human Pose Estimation in VideoYu Cheng, Bo Yang, Bo Wang, Wending Yan 等ICCV 2019 · 被引用 223 次
相关 Paper
- Combining Detection and Tracking for Human Pose Estimation in VideosManchen Wang, Joseph Tighe, Davide ModoloCVPR 2020
- 3D Human Pose Estimation Using Spatio-Temporal Networks with Explicit Occlusion TrainingYu Cheng, Bo Yang, Bo Wang, Robby T. TanAAAI 2020 · 被引用 145 次
- TwinPose: Person-Specific Subspaces for Multi-View 3D Pose EstimationWenwu Yang, Tianyi He, Jiwei Ding, Xun Wang 等SIGGRAPH 2026
- Simple Pose: Rethinking and Improving a Bottom-up Approach for Multi-Person Pose EstimationJia Li, Wen Su, Zengfu WangAAAI 2020 · 被引用 104 次
- Probabilistic Monocular 3D Human Pose Estimation with Normalizing FlowsTom Wehrbein, Marco Rudolph, Bodo Rosenhahn, Bastian WandtICCV 2021 · 被引用 147 次
