TexturePose: Supervising Human Mesh Estimation With Texture Consistency
Georgios Pavlakos, Nikos Kolotouros, Kostas Daniilidis
Abstract
This work addresses the problem of model-based human pose estimation. Recent approaches have made significant progress towards regressing the parameters of parametric human body models directly from images. Because of the absence of images with 3D shape ground truth, relevant approaches rely on 2D annotations or sophisticated architecture designs. In this work, we advocate that there are more cues we can leverage, which are available for free in natural images, i.e., without getting more annotations, or modifying the network architecture. We propose a natural form of supervision, that capitalizes on the appearance constancy of a person among different frames (or viewpoints). This seemingly insignificant and often overlooked cue goes a long way for model-based pose estimation. The parametric model we employ allows us to compute a texture map for each frame. Assuming that the texture of the person does not change dramatically between frames, we can apply a novel texture consistency loss, which enforces that each point in the texture map has the same texture value across all frames. Since the texture is transferred in this common texture map space, no camera motion computation is necessary, or even an assumption of smoothness among frames. This makes our proposed supervision applicable in a variety of settings, ranging from monocular video, to multi-view images. We benchmark our approach against strong baselines that require the same or even more annotations that we do and we consistently outperform them. Simultaneously, we achieve state-of-the-art results among model-based pose estimation approaches in different benchmarks. The project website with videos, results, and code can be found at https: //seas.upenn.edu/ ˜pavlakos/projects/texturepose.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a7ed0d76-5272-4df5-abb0-619e3e4fb909Cited by top-tier papers38
- PyMAF: 3D Human Pose and Shape Regression with Pyramidal Mesh Alignment Feedback LoopHongwen Zhang, Yating Tian, Xinchi Zhou, Wanli Ouyang et al.ICCV 2021 · 376 citations
- Monocular, One-stage, Regression of Multiple 3D PeopleYu Sun, Qian Bao, Wu Liu, Yili Fu et al.ICCV 2021 · 335 citations
- Geo-PIFu: Geometry and Pixel Aligned Implicit Functions for Single-view Human ReconstructionTong He, John P. Collomosse, Hailin Jin, Stefano SoattoNeurIPS 2020 · 203 citations
- Probabilistic Modeling for Human Mesh RecoveryNikos Kolotouros, Georgios Pavlakos, Dinesh Jayaraman, Kostas DaniilidisICCV 2021 · 201 citations
- Putting People in their Place: Monocular Regression of 3D People in DepthYu Sun, Wu Liu, Qian Bao, Yili Fu et al.CVPR 2022 · 152 citations
Related papers
- Human Parsing Based Texture Transfer from Single Image to 3D Human via Cross-View ConsistencyFang Zhao, Shengcai Liao, Kaihao Zhang, Ling ShaoNeurIPS 2020 · 22 citations
- CanonPose: Self-Supervised Monocular 3D Human Pose Estimation in the WildBastian Wandt, Marco Rudolph, Petrissa Zell, Helge Rhodin et al.CVPR 2021
- Weakly-Supervised 3D Human Pose Learning via Multi-View Images in the WildUmar Iqbal, Pavlo Molchanov, Jan KautzCVPR 2020
- Multiview-Consistent Semi-Supervised Learning for 3D Human Pose EstimationRahul Mitra, Nitesh B. Gundavarapu, Abhishek Sharma, Arjun JainCVPR 2020
- Geometry-Driven Self-Supervised Method for 3D Human Pose EstimationYang Li, Kan Li, Shuai Jiang, Ziyue Zhang et al.AAAI 2020 · 40 citations
