On the Calibration of Human Pose Estimation
Kerui Gu, Rongyu Chen, Xuanlong Yu, Angela Yao
摘要
Most 2D human pose estimation frameworks estimate keypoint confidence in an ad-hoc manner, using heuristics such as the maximum value of heatmaps. The confidence is part of the evaluation scheme, e.g., AP for the MSCOCO dataset, yet has been largely overlooked in the development of state-of-the-art methods. This paper takes the first steps in addressing miscalibration in pose estimation. From a calibration point of view, the confidence should be aligned with the pose accuracy. In practice, existing methods are poorly calibrated. We show, through theoretical analysis, why a miscalibration gap exists and how to narrow the gap. Simply predicting the instance size and adjusting the confidence function gives considerable AP improvements. Given the black-box nature of deep neural networks, however, it is not possible to fully close this gap with only closed-form adjustments. As such, we go one step further and learn network-specific adjustments by enforcing consistency between confidence and pose accuracy. Our proposed Calibrated ConfidenceNet (CCNet) is a light-weight post-hoc addition that improves AP by up to 1.4% on offthe-shelf pose estimation frameworks. Applied to the downstream task of mesh recovery, CCNet facilitates an additional 1.0mm decrease in 3D keypoint error. Project page: comp.nus.edu.sg/keruigu/calibrate pose/project.html.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- KITRO: Refining Human Mesh by 2D Clues and Kinematic-tree RotationFengyuan Yang, Kerui Gu, Angela YaoCVPR 2024 · 被引用 6 次
- Semantics-aware Test-time Adaptation for 3D Human Pose EstimationQiuxia Lin, Rongyu Chen, Kerui Gu, Angela YaoICML 2025
- ExtPose: Robust and Coherent Pose Estimation by Extending ViTsRongyu Chen, Li'an Zhuo, Linlin Yang, Qi Wang 等ICML 2025
- ProbPose: A Probabilistic Approach to 2D Human Pose EstimationMiroslav Purkrábek, Jiri MatasCVPR 2025
- Detection, Pose Estimation and Segmentation for Multiple Bodies: Closing the Virtuous CircleMiroslav Purkrábek, Jiri MatasICCV 2025
它引用的顶会 Paper17
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 被引用 6,759 次
- Learning to Reconstruct 3D Human Pose and Shape via Model-Fitting in the LoopNikos Kolotouros, Georgios Pavlakos, Michael J. Black, Kostas DaniilidisICCV 2019 · 被引用 1,139 次
- ViTPose: Simple Vision Transformer Baselines for Human Pose EstimationYufei Xu, Jing Zhang, Qiming Zhang, Dacheng TaoNeurIPS 2022 · 被引用 1,105 次
- Deep Evidential RegressionAlexander Amini, Wilko Schwarting, Ava Soleimany, Daniela RusNeurIPS 2020 · 被引用 777 次
- Human Pose Regression with Residual Log-likelihood EstimationJiefeng Li, Siyuan Bian, Ailing Zeng, Can Wang 等ICCV 2021 · 被引用 286 次
相关 Paper
- Learning Quality-Aware Representation for Multi-Person Pose RegressionYabo Xiao, Dongdong Yu, Xiaojuan Wang, Lei Jin 等AAAI 2022 · 被引用 17 次
- HigherHRNet: Scale-Aware Representation Learning for Bottom-Up Human Pose EstimationBowen Cheng, Bin Xiao, Jingdong Wang, Honghui Shi 等CVPR 2020
- Anchor Loss: Modulating Loss Scale Based on Prediction DifficultySerim Ryou, Seong-Gyun Jeong, Pietro PeronaICCV 2019 · 被引用 46 次
- MonoLoco: Monocular 3D Pedestrian Localization and Uncertainty EstimationLorenzo Bertoni, Sven Kreiss, Alexandre AlahiICCV 2019 · 被引用 125 次
- DecenterNet: Bottom-Up Human Pose Estimation Via Decentralized Pose RepresentationTao Wang, Lei Jin, Zhang Wang, Xiaojin Fan 等ACM MM 2023 · 被引用 14 次
