InclusiveVidPose: Bridging the Pose Estimation Gap for Individuals with Limb Deficiencies in Video-Based Motion
Heming Du, Jiaying Ying, Sen Wang, Xue Li, Kaihao Zhang, Xin Yu
摘要
Approximately 445.2 million individuals worldwide are living with traumatic amputations, and an estimated 31.64 million children aged 0–14 have congenital limb differences, yet they remain largely underrepresented in human pose estimation (HPE) research. Accurate HPE could significantly benefit this population in applications, such as rehabilitation monitoring and health assessment. However, the existing HPE datasets and methods assume that humans possess a full complement of upper and lower extremities and fail to model missing or altered limbs. As a result, people with limb deficiencies remain largely underrepresented, and current models cannot generalize to their unique anatomies or predict absent joints. To bridge this gap, we introduce InclusiveVidPose Dataset, the first video-based large-scale HPE dataset specific for individuals with limb deficiencies. We collect 313 videos, totaling 327k frames, and covering nearly 400 individuals with amputations, congenital limb differences, and prosthetic limbs. We adopt 8 extra keypoints at each residual limb end to capture individual anatomical variations. Under the guidance of an internationally accredited para-athletics classifier, we annotate each frame with pose keypoints, segmentation masks, bounding boxes, tracking IDs, and per-limb prosthesis status. Experiments on InclusiveVidPose highlight the limitations of the existing HPE models for individuals with limb deficiencies. We introduce a new evaluation metric, Limb-specific Confidence Consistency (LiCC), which assesses the consistency of pose estimations between residual and intact limb keypoints. We also provide a rigorous benchmark for evaluating inclusive and robust pose estimation algorithms, demonstrating that our dataset poses significant challenges. We hope InclusiveVidPose spur research toward methods that fairly and accurately serve all body types. The project website is available at: InclusiveVidPose.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper12
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- ViTPose: Simple Vision Transformer Baselines for Human Pose EstimationYufei Xu, Jing Zhang, Qiming Zhang, Dacheng TaoNeurIPS 2022 · 被引用 1,105 次
- GoPose: 3D Human Pose Estimation Using WiFiYili Ren, Zi Wang, Yichao Wang, Sheng Tan 等UbiComp 2022 · 被引用 100 次
- PoseTrack21: A Dataset for Person Search, Multi-Object Tracking and Multi-Person Pose TrackingAndreas Doering, Di Chen, Shanshan Zhang, Bernt Schiele 等CVPR 2022 · 被引用 47 次
- WheelPose: Data Synthesis Techniques to Improve Pose Estimation Performance on Wheelchair UsersWilliam Huang, Sam Ghahremani, Siyou Pei, Yang ZhangCHI 2024 · 被引用 12 次
相关 Paper
- LDPose: Towards Inclusive Human Pose Estimation for Limb-Deficient Individuals in the WildJiaying Ying, Heming Du, Kaihao Zhang, Lincheng Li 等ICCV 2025 · 被引用 3 次
- IVQA-LD: Inclusive Multimodal Understanding for Population with Limb DeficiencyYan Ke, Xin Shen, Jiaying Ying, Xin Li 等ICML 2026
- ProGait: A Multi-Purpose Video Dataset and Benchmark for Transfemoral Prosthesis UsersXiangyu Yin, Boyuan Yang, Weichen Liu, Qiyao Xue 等ICCV 2025 · 被引用 5 次
- AJAHR: Amputated Joint Aware 3D Human Mesh RecoveryHyunjin Cho, Giyun Choi, Jongwon ChoiICCV 2025 · 被引用 1 次
- Cross-View Tracking for Multi-Human 3D Pose Estimation at Over 100 FPSLong Chen, Haizhou Ai, Rui Chen, Zijie Zhuang 等CVPR 2020
