Bottom-Up Human Pose Estimation via Disentangled Keypoint Regression
Zigang Geng, Ke Sun, Bin Xiao, Zhaoxiang Zhang, Jingdong Wang
Abstract
In this paper, we are interested in the bottom-up paradigm of estimating human poses from an image. We study the dense keypoint regression framework that is previously inferior to the keypoint detection and grouping framework. Our motivation is that regressing keypoint positions accurately needs to learn representations that focus on the keypoint regions. We present a simple yet effective approach, named disentangled keypoint regression (DEKR). We adopt adaptive convolutions through pixel-wise spatial transformer to activate the pixels in the keypoint regions and accordingly learn representations from them. We use a multi-branch structure for separate regression: each branch learns a representation with dedicated adaptive convolutions and regresses one keypoint. The resulting disentangled representations are able to attend to the keypoint regions, respectively, and thus the keypoint regression is spatially more accurate. We empirically show that the proposed direct regression method outperforms keypoint detection and grouping methods and achieves superior bottom-up pose estimation results on two benchmark datasets, COCO and Crowd-Pose. The code and models are available at https: //github.com/HRNet/DEKR .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c070638c-18e7-46e3-915b-ba15edb9dae6Cited by top-tier papers52
- Conditional DETR for Fast Training ConvergenceDepu Meng, Xiaokang Chen, Zejia Fan, Gang Zeng et al.ICCV 2021 · 974 citations
- End-to-End Multi-Person Pose Estimation with TransformersDahu Shi, Xing Wei, Liangqi Li, Ye Ren et al.CVPR 2022 · 147 citations
- Lite Pose: Efficient Architecture Design for 2D Human Pose EstimationYihan Wang, Muyang Li, Han Cai, Wei-Ming Chen et al.CVPR 2022 · 117 citations
- Contextual Instance Decoupling for Robust Multi-Person Pose EstimationDongkai Wang, Shiliang ZhangCVPR 2022 · 73 citations
- RTMO: Towards High-Performance One-Stage Real-Time Multi-Person Pose EstimationPeng Lu, Tao Jiang, Yining Li, Xiangtai Li et al.CVPR 2024 · 66 citations
Builds on9
- CenterNet: Keypoint Triplets for Object DetectionKaiwen Duan, Song Bai, Lingxi Xie, Honggang Qi et al.ICCV 2019 · 3,348 citations
- Single-Stage Multi-Person Pose MachinesXuecheng Nie, Jiashi Feng, Jianfeng Zhang, Shuicheng YanICCV 2019 · 246 citations
- Simple Pose: Rethinking and Improving a Bottom-up Approach for Multi-Person Pose EstimationJia Li, Wen Su, Zengfu WangAAAI 2020 · 104 citations
- Mixture Dense Regression for Object Detection and Human Pose EstimationAli Varamesh, Tinne TuytelaarsCVPR 2020
- S3VAE: Self-Supervised Sequential VAE for Representation Disentanglement and Data GenerationYizhe Zhu, Martin Renqiang Min, Asim Kadav, Hans Peter GrafCVPR 2020
Related papers
- The Center of Attention: Center-Keypoint Grouping via Attention for Multi-Person Pose EstimationGuillem Brasó, Nikita Kister, Laura Leal-TaixéICCV 2021 · 50 citations
- Semantic-aware Transfer with Instance-adaptive Parsing for Crowded Scenes Pose EstimationXuanhan Wang, Lianli Gao, Yan Dai, Yixuan Zhou et al.ACM MM 2021 · 14 citations
- QueryPose: Sparse Multi-Person Pose Regression via Spatial-Aware Part-Level QueryYabo Xiao, Kai Su, Xiaojuan Wang, Dongdong Yu et al.NeurIPS 2022 · 32 citations
- Group Pose: A Simple Baseline for End-to-End Multi-person Pose EstimationHuan Liu, Qiang Chen, Zichang Tan, Jiang-Jiang Liu et al.ICCV 2023 · 50 citations
- AdaptivePose: Human Parts as Adaptive PointsYabo Xiao, Xiaojuan Wang, Dongdong Yu, Guoli Wang et al.AAAI 2022 · 25 citations
