Unsupervised Human Pose Estimation Through Transforming Shape Templates
Luca Schmidtke, Athanasios Vlontzos, Simon Ellershaw, Anna Lukens, Tomoki Arichi, Bernhard Kainz
Abstract
Human pose estimation is a major computer vision problem with applications ranging from augmented reality and video capture to surveillance and movement tracking. In the medical context, the latter may be an important biomarker for neurological impairments in infants. Whilst many methods exist, their application has been limited by the need for well annotated large datasets and the inability to generalize to humans of different shapes and body compositions, e.g. children and infants. In this paper we present a novel method for learning pose estimators for human adults and infants in an unsupervised fashion. We approach this as a learnable template matching problem facilitated by deep feature extractors. Human-interpretable landmarks are estimated by transforming a template consisting of predefined body parts that are characterized by 2D Gaussian distributions. Enforcing a connectivity prior guides our model to meaningful human shape representations. We demonstrate the effectiveness of our approach on two different datasets including adults and infants. Project
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2f388ed9-3fdf-4969-96ee-0d2a786161d1Cited by top-tier papers16
- Temporal Feature Alignment and Mutual Information Maximization for Video-Based Human Pose EstimationZhenguang Liu, Runyang Feng, Haoming Chen, Shuang Wu et al.CVPR 2022 · 76 citations
- DiffPose: SpatioTemporal Diffusion Model for Video-Based Human Pose EstimationRunyang Feng, Yixing Gao, Tze Ho Elden Tse, Xueqing Ma et al.ICCV 2023 · 46 citations
- Watch It Move: Unsupervised Discovery of 3D Joints for Re-Posing of Articulated ObjectsAtsuhiro Noguchi, Umar Iqbal, Jonathan Tremblay, Tatsuya Harada et al.CVPR 2022 · 32 citations
- AutoLink: Self-supervised Learning of Human Skeletons and Object Outlines by Linking KeypointsXingzhe He, Bastian Wandt, Helge RhodinNeurIPS 2022 · 28 citations
- Self-Supervised Keypoint Discovery in Behavioral VideosJennifer J. Sun, Serim Ryou, Roni H. Goldshmid, Brandon Weissbourd et al.CVPR 2022 · 24 citations
Builds on3
- Unsupervised Learning of Landmarks by Descriptor Vector ExchangeJames Thewlis, Samuel Albanie, Hakan Bilen, Andrea VedaldiICCV 2019 · 70 citations
- Self-Supervised Learning of Interpretable Keypoints From Unlabelled VideosTomas Jakab, Ankush Gupta, Hakan Bilen, Andrea VedaldiCVPR 2020
- Self-Supervised 3D Human Pose Estimation via Part Guided Novel Image SynthesisJogendra Nath Kundu, Siddharth Seth, Varun Jampani, Mugalodi Rakesh et al.CVPR 2020
Related papers
- Pose Prior Learner: Unsupervised Categorical Prior Learning for Pose EstimationZiyu Wang, Shuangpeng Han, Mengmi ZhangICLR 2026 · 3 citations
- LEAP: Learning Articulated Occupancy of PeopleMarko Mihajlovic, Yan Zhang, Michael J. Black, Siyu TangCVPR 2021
- 3D Human Mesh Estimation from Virtual MarkersXiaoxuan Ma, Jiajun Su, Chunyu Wang, Wentao Zhu et al.CVPR 2023
- 3D-Aware Neural Body Fitting for Occlusion Robust 3D Human Pose EstimationYi Zhang, Pengliang Ji, Angtian Wang, Jieru Mei et al.ICCV 2023 · 44 citations
- COAP: Compositional Articulated Occupancy of PeopleMarko Mihajlovic, Shunsuke Saito, Aayush Bansal, Michael Zollhöfer et al.CVPR 2022 · 45 citations
