DecenterNet: Bottom-Up Human Pose Estimation Via Decentralized Pose Representation
Tao Wang, Lei Jin, Zhang Wang, Xiaojin Fan, Yu Cheng, Yinglei Teng, Junliang Xing, Jian Zhao
Abstract
Multi-person pose estimation in crowded scenes remains a very challenging task. This paper finds that most previous methods fail to estimate or group visible keypoints in crowded scenes rather than reasoning invisible keypoints. We thus categorize the crowded scenes into entanglement and occlusion based on the visibility of human parts and observe that entanglement is a significant problem in crowded scenes. With this observation, we propose DecenterNet, an end-to-end deep architecture to perform robust and efficient pose estimation in crowded scenes. Within DecenterNet, we introduce a decentralized pose representation that uses all visible keypoints as the root points to represent human poses, which is more robust in the entanglement area. We also propose a decoupled pose assessment mechanism, which introduces a location map to adaptively select optimal poses in the offset map. In addition, we have constructed a new dataset named SkatingPose, containing more entangled scenes. The proposed DecenterNet surpasses the best method on SkatingPose by 1.8 AP. Furthermore, DecenterNet obtains 71.2 AP and 71.4 AP on the COCO and CrowdPose datasets, respectively, demonstrating the superiority of our method. We will release our source code, trained models, and dataset to facilitate further studies in this research direction. Our code and dataset are available in https://github.com/InvertedForest/DecenterNet.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 83a82688-ce9b-443e-8b8c-a7f99be878f3Cited by top-tier papers2
- OpenMoCap: Rethinking Optical Motion Capture under Real-world OcclusionChen Qian, Danyang Li, Xinran Yu, Zheng Yang et al.ACM MM 2025 · 1 citation
- SynSP: Synergy of Smoothness and Precision in Pose Sequences RefinementTao Wang, Lei Jin, Zheng Wang, Jianshu Li et al.CVPR 2024
Related papers
- Bottom-Up Human Pose Estimation via Disentangled Keypoint RegressionZigang Geng, Ke Sun, Bin Xiao, Zhaoxiang Zhang et al.CVPR 2021
- Contextual Instance Decoupling for Robust Multi-Person Pose EstimationDongkai Wang, Shiliang ZhangCVPR 2022 · 73 citations
- Robust Pose Estimation in Crowded Scenes with Direct Pose-Level InferenceDongkai Wang, Shiliang Zhang, Gang HuaNeurIPS 2021 · 36 citations
- DiffusionRegPose: Enhancing Multi-Person Pose Estimation Using a Diffusion-Based End-to-End Regression ApproachDayi Tan, Hansheng Chen, Wei Tian, Lu XiongCVPR 2024 · 6 citations
- Learning to Estimate Robust 3D Human Mesh from In-the-Wild Crowded ScenesHongsuk Choi, Gyeongsik Moon, JoonKyu Park, Kyoung Mu LeeCVPR 2022 · 92 citations
