Single-Stage is Enough: Multi-Person Absolute 3D Pose Estimation
Lei Jin, Chenyang Xu, Xiaojuan Wang, Yabo Xiao, Yandong Guo, Xuecheng Nie, Jian Zhao
Abstract
The existing multi-person absolute 3D pose estimation methods are mainly based on two-stage paradigm, i.e., top-down or bottom-up, leading to redundant pipelines with high computation cost. We argue that it is more desirable to simplify such two-stage paradigm to a single-stage one to promote both efficiency and performance. To this end, we present an efficient single-stage solution, Decoupled Regression Model (DRM), with three distinct novelties. First, DRM introduces a new decoupled representation for 3D pose, which expresses the 2D pose in image plane and depth information of each 3D human instance via 2D center point (center of visible keypoints) and root point (denoted as pelvis), respectively. Second, to learn better feature representation for the human depth regression, DRM introduces a 2D Pose-guided Depth Query Module (PDQM) to extract the features in 2D pose regression branch, enabling the depth regression branch to perceive the scale information of instances. Third, DRM leverages a Decoupled Absolute Pose Loss (DAPL) to facilitate the absolute root depth and root-relative depth estimation, thus improving the accuracy of absolute 3D pose. Comprehensive experiments on challenging benchmarks including MuPoTS-3D and Panoptic clearly verify the superiority of our framework, which outperforms the state-of-the-art bottom-up absolute 3D pose estimation methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9c040ff1-bd11-4de7-80de-123b4ee321a8Cited by top-tier papers4
- Reconstructing Groups of People with Hypergraph Relational ReasoningBuzhen Huang, Jingyi Ju, Zhihao Li, Yangang WangICCV 2023 · 21 citations
- Towards Robust and Smooth 3D Multi-Person Pose Estimation from Monocular Videos in the WildSungchan Park, Eunyi You, Inhoe Lee, Joonseok LeeICCV 2023 · 16 citations
- SynSP: Synergy of Smoothness and Precision in Pose Sequences RefinementTao Wang, Lei Jin, Zheng Wang, Jianshu Li et al.CVPR 2024
- Towards Stable Human Pose Estimation via Cross-View Fusion and Foot StabilizationLi'an Zhuo, Jian Cao, Qi Wang, Bang Zhang et al.CVPR 2023
Builds on11
- Camera Distance-Aware Top-Down Approach for 3D Multi-Person Pose Estimation From a Single RGB ImageGyeongsik Moon, Ju Yong Chang, Kyoung Mu LeeICCV 2019 · 368 citations
- XNect: real-time multi-person 3D motion capture with a single RGB cameraDushyant Mehta, Oleksandr Sotnychenko, Franziska Mueller, Weipeng Xu et al.SIGGRAPH 2020 · 267 citations
- Single-Stage Multi-Person Pose MachinesXuecheng Nie, Jiashi Feng, Jianfeng Zhang, Shuicheng YanICCV 2019 · 246 citations
- 3D Human Pose Estimation Using Spatio-Temporal Networks with Explicit Occlusion TrainingYu Cheng, Bo Yang, Bo Wang, Robby T. TanAAAI 2020 · 145 citations
- AdaptivePose: Human Parts as Adaptive PointsYabo Xiao, Xiaojuan Wang, Dongdong Yu, Guoli Wang et al.AAAI 2022 · 25 citations
Related papers
- Distribution-Aware Single-Stage Models for Multi-Person 3D Pose EstimationZitian Wang, Xuecheng Nie, Xiaochao Qu, Yunpeng Chen et al.CVPR 2022 · 44 citations
- Body Meshes as PointsJianfeng Zhang, Dongdong Yu, Jun Hao Liew, Xuecheng Nie et al.CVPR 2021
- DGG-HMR: Multi-Person Human Mesh Recovery with Depth-Guided Geometric AnchoringYanjie Li, Le Hui, Yali Peng, Shigang LiuICML 2026
- Mutual Adaptive Reasoning for Monocular 3D Multi-Person Pose EstimationJuze Zhang, Jingya Wang, Ye Shi, Fei Gao et al.ACM MM 2022 · 15 citations
- Single-Stage Multi-human Parsing via Point Sets and Center-based OffsetsJiaming Chu, Lei Jin, Xiaojin Fan, Yinglei Teng et al.ACM MM 2023 · 14 citations
