Invariant Teacher and Equivariant Student for Unsupervised 3D Human Pose Estimation
Chenxin Xu, Siheng Chen, Maosen Li, Ya Zhang
Abstract
We propose a novel method based on teacher-student learning framework for 3D human pose estimation without any 3D annotation or side information. To solve this unsupervised-learning problem, the teacher network adopts pose-dictionary-based modeling for regularization to estimate a physically plausible 3D pose. To handle the decomposition ambiguity in the teacher network, we propose a cycle-consistent architecture promoting a 3D rotation-invariant property to train the teacher network. To further improve the estimation accuracy, the student network adopts a novel graph convolution network for flexibility to directly estimate the 3D coordinates. Another cycle-consistent architecture promoting 3D rotation-equivariant property is adopted to exploit geometry consistency, together with knowledge distillation from the teacher network to improve the pose estimation performance. We conduct extensive experiments on Human3.6M and MPI-INF-3DHP. Our method reduces the 3D joint prediction error by 11.4% compared to state-of-the-art unsupervised methods and also outperforms many weakly-supervised methods that use side information on Human3.6M. Code will be available at https://github.com/sjtuxcx/ITES.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3f04bf21-fea5-4e4b-8482-854465c472c4Cited by top-tier papers5
- Auxiliary Tasks Benefit 3D Skeleton-based Human Motion PredictionChenxin Xu, Robby T. Tan, Yuhong Tan, Siheng Chen et al.ICCV 2023 · 35 citations
- Pose-Transformed Equivariant Network for 3D Point Trajectory PredictionRuixuan Yu, Jian SunCVPR 2024 · 2 citations
- Deep Non-Rigid Structure-from-Motion Revisited: Canonicalization and Sequence ModelingHui Deng, Jiawei Shi, Zhen Qin, Yiran Zhong et al.AAAI 2025 · 1 citation
- Unsupervised 3D Structure Inference from Category-Specific Image CollectionsWeikang Wang, Dongliang Cao, Florian BernardCVPR 2024
- EqMotion: Equivariant Multi-Agent Motion Prediction with Invariant Interaction ReasoningChenxin Xu, Robby T. Tan, Yuhong Tan, Siheng Chen et al.CVPR 2023
Builds on5
- Learning to Reconstruct 3D Human Pose and Shape via Model-Fitting in the LoopNikos Kolotouros, Georgios Pavlakos, Michael J. Black, Kostas DaniilidisICCV 2019 · 1,139 citations
- Optimizing Network Structure for 3D Human Pose EstimationHai Ci, Chunyu Wang, Xiaoxuan Ma, Yizhou WangICCV 2019 · 267 citations
- Distill Knowledge From NRSfM for Weakly Supervised 3D Pose LearningChaoyang Wang, Chen Kong, Simon LuceyICCV 2019 · 52 citations
- Geometry-Driven Self-Supervised Method for 3D Human Pose EstimationYang Li, Kan Li, Shuai Jiang, Ziyue Zhang et al.AAAI 2020 · 40 citations
- Dynamic Multiscale Graph Neural Networks for 3D Skeleton Based Human Motion PredictionMaosen Li, Siheng Chen, Yangheng Zhao, Ya Zhang et al.CVPR 2020
Related papers
- Weakly-Supervised 3D Human Pose Learning via Multi-View Images in the WildUmar Iqbal, Pavlo Molchanov, Jan KautzCVPR 2020
- Kinematic-Structure-Preserved Representation for Unsupervised 3D Human Pose EstimationJogendra Nath Kundu, Siddharth Seth, Rahul M. V., Mugalodi Rakesh et al.AAAI 2020 · 57 citations
- Chained Representation Cycling: Learning to Estimate 3D Human Pose and Shape by Cycling Between RepresentationsNadine Rueegg, Christoph Lassner, Michael J. Black, Konrad SchindlerAAAI 2020 · 25 citations
- Multiview-Consistent Semi-Supervised Learning for 3D Human Pose EstimationRahul Mitra, Nitesh B. Gundavarapu, Abhishek Sharma, Arjun JainCVPR 2020
- MAPConNet: Self-supervised 3D Pose Transfer with Mesh and Point Contrastive LearningJiaze Sun, Zhixiang Chen, Tae-Kyun KimICCV 2023 · 2 citations
