Distill Knowledge From NRSfM for Weakly Supervised 3D Pose Learning
Chaoyang Wang, Chen Kong, Simon Lucey
Abstract
We propose to learn a 3D pose estimator by distilling knowledge from Non-Rigid Structure from Motion (NRSfM). Our method uses solely 2D landmark annotations. No 3D data, multi-view/temporal footage, or object specific prior is required. This alleviates the data bottleneck, which is one of the major concern for supervised methods. The challenge for using NRSfM as teacher is that they often make poor depth reconstruction when the 2D projections have strong ambiguity. Directly using those wrong depth as hard target would negatively impact the student. Instead, we propose a novel loss that ties depth prediction to the cost function used in NRSfM. This gives the student pose estimator freedom to reduce depth error by associating with image features. Validated on H3.6M dataset, our learned 3D pose estimation network achieves more accurate reconstruction compared to NRSfM methods. It also outperforms other weakly supervised methods, in spite of using significantly less supervision.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0c3c9319-ee10-40f1-b149-76c581d831efCited by top-tier papers11
- SDF-SRN: Learning Signed Distance 3D Object Reconstruction from Static ImagesChen-Hsuan Lin, Chaoyang Wang, Simon LuceyNeurIPS 2020 · 125 citations
- Online Knowledge Distillation for Efficient Pose EstimationZheng Li, Jingwen Ye, Mingli Song, Ying Huang et al.ICCV 2021 · 123 citations
- Estimating Egocentric 3D Human Pose in the Wild with External Weak SupervisionJian Wang, Lingjie Liu, Weipeng Xu, Kripasindhu Sarkar et al.CVPR 2022 · 33 citations
- Invariant Teacher and Equivariant Student for Unsupervised 3D Human Pose EstimationChenxin Xu, Siheng Chen, Maosen Li, Ya ZhangAAAI 2021 · 20 citations
- Deductive Learning for Weakly-Supervised 3D Human Pose Estimation via Uncalibrated CamerasXipeng Chen, Pengxu Wei, Liang LinAAAI 2021 · 14 citations
Related papers
- Geometry-Driven Self-Supervised Method for 3D Human Pose EstimationYang Li, Kan Li, Shuai Jiang, Ziyue Zhang et al.AAAI 2020 · 40 citations
- Towards Alleviating the Modeling Ambiguity of Unsupervised Monocular 3D Human Pose EstimationZhenbo Yu, Bingbing Ni, Jingwei Xu, Junjie Wang et al.ICCV 2021 · 39 citations
- Deep Non-Rigid Structure From MotionChen Kong, Simon LuceyICCV 2019 · 72 citations
- PAUL: Procrustean Autoencoder for Unsupervised LiftingChaoyang Wang, Simon LuceyCVPR 2021
- Deep Non-Rigid Structure-from-Motion Revisited: Canonicalization and Sequence ModelingHui Deng, Jiawei Shi, Zhen Qin, Yiran Zhong et al.AAAI 2025 · 1 citation
