InclusiveVidPose: Bridging the Pose Estimation Gap for Individuals with Limb Deficiencies in Video-Based Motion
Heming Du, Jiaying Ying, Sen Wang, Xue Li, Kaihao Zhang, Xin Yu
Abstract
Approximately 445.2 million individuals worldwide are living with traumatic amputations, and an estimated 31.64 million children aged 0–14 have congenital limb differences, yet they remain largely underrepresented in human pose estimation (HPE) research. Accurate HPE could significantly benefit this population in applications, such as rehabilitation monitoring and health assessment. However, the existing HPE datasets and methods assume that humans possess a full complement of upper and lower extremities and fail to model missing or altered limbs. As a result, people with limb deficiencies remain largely underrepresented, and current models cannot generalize to their unique anatomies or predict absent joints. To bridge this gap, we introduce InclusiveVidPose Dataset, the first video-based large-scale HPE dataset specific for individuals with limb deficiencies. We collect 313 videos, totaling 327k frames, and covering nearly 400 individuals with amputations, congenital limb differences, and prosthetic limbs. We adopt 8 extra keypoints at each residual limb end to capture individual anatomical variations. Under the guidance of an internationally accredited para-athletics classifier, we annotate each frame with pose keypoints, segmentation masks, bounding boxes, tracking IDs, and per-limb prosthesis status. Experiments on InclusiveVidPose highlight the limitations of the existing HPE models for individuals with limb deficiencies. We introduce a new evaluation metric, Limb-specific Confidence Consistency (LiCC), which assesses the consistency of pose estimations between residual and intact limb keypoints. We also provide a rigorous benchmark for evaluating inclusive and robust pose estimation algorithms, demonstrating that our dataset poses significant challenges. We hope InclusiveVidPose spur research toward methods that fairly and accurately serve all body types. The project website is available at: InclusiveVidPose.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a2f7fb70-ee25-44a9-b47d-732d369b8aeaBuilds on12
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- ViTPose: Simple Vision Transformer Baselines for Human Pose EstimationYufei Xu, Jing Zhang, Qiming Zhang, Dacheng TaoNeurIPS 2022 · 1,105 citations
- GoPose: 3D Human Pose Estimation Using WiFiYili Ren, Zi Wang, Yichao Wang, Sheng Tan et al.UbiComp 2022 · 100 citations
- PoseTrack21: A Dataset for Person Search, Multi-Object Tracking and Multi-Person Pose TrackingAndreas Doering, Di Chen, Shanshan Zhang, Bernt Schiele et al.CVPR 2022 · 47 citations
- WheelPose: Data Synthesis Techniques to Improve Pose Estimation Performance on Wheelchair UsersWilliam Huang, Sam Ghahremani, Siyou Pei, Yang ZhangCHI 2024 · 12 citations
Related papers
- LDPose: Towards Inclusive Human Pose Estimation for Limb-Deficient Individuals in the WildJiaying Ying, Heming Du, Kaihao Zhang, Lincheng Li et al.ICCV 2025 · 3 citations
- IVQA-LD: Inclusive Multimodal Understanding for Population with Limb DeficiencyYan Ke, Xin Shen, Jiaying Ying, Xin Li et al.ICML 2026
- ProGait: A Multi-Purpose Video Dataset and Benchmark for Transfemoral Prosthesis UsersXiangyu Yin, Boyuan Yang, Weichen Liu, Qiyao Xue et al.ICCV 2025 · 5 citations
- AJAHR: Amputated Joint Aware 3D Human Mesh RecoveryHyunjin Cho, Giyun Choi, Jongwon ChoiICCV 2025 · 1 citation
- Cross-View Tracking for Multi-Human 3D Pose Estimation at Over 100 FPSLong Chen, Haizhou Ai, Rui Chen, Zijie Zhuang et al.CVPR 2020
