DenseRaC: Joint 3D Pose and Shape Estimation by Dense Render-and-Compare
Yuanlu Xu, Song-Chun Zhu, Tony Tung
Abstract
We present DenseRaC, a novel end-to-end framework for jointly estimating 3D human pose and body shape from a monocular RGB image. Our two-step framework takes the body pixel-to-surface correspondence map (i.e., IUV map) as proxy representation and then performs estimation of parameterized human pose and shape. Specifically, given an estimated IUV map, we develop a deep neural network optimizing 3D body reconstruction losses and further integrating a render-and-compare scheme to minimize differences between the input and the rendered output, i.e., dense body landmarks, body part masks, and adversarial priors. To boost learning, we further construct a large-scale synthetic dataset (MOCA) utilizing web-crawled Mocap sequences, 3D scans and animations. The generated data covers diversified camera views, human actions and body shapes, and is paired with full ground truth. Our model jointly learns to represent the 3D human body from hybrid datasets, mitigating the problem of unpaired training data. Our experiments show that DenseRaC obtains superior performance against state of the art on public benchmarks of various human-related tasks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers61
- Humans in 4D: Reconstructing and Tracking Humans with TransformersShubham Goel, Georgios Pavlakos, Jathushan Rajasegaran, Angjoo Kanazawa et al.ICCV 2023 · 390 citations
- PyMAF: 3D Human Pose and Shape Regression with Pyramidal Mesh Alignment Feedback LoopHongwen Zhang, Yating Tian, Xinchi Zhou, Wanli Ouyang et al.ICCV 2021 · 376 citations
- Monocular, One-stage, Regression of Multiple 3D PeopleYu Sun, Qian Bao, Wu Liu, Yili Fu et al.ICCV 2021 · 335 citations
- MotionBERT: A Unified Perspective on Learning Human Motion RepresentationsWentao Zhu, Xiaoxuan Ma, Zhaoyang Liu, Libin Liu et al.ICCV 2023 · 322 citations
- ARCH++: Animation-Ready Clothed Human Reconstruction RevisitedTong He, Yuanlu Xu, Shunsuke Saito, Stefano Soatto et al.ICCV 2021 · 233 citations
Related papers
- UltraPose: Synthesizing Dense Pose with 1 Billion Points by Human-body Decoupling 3D ModelHaonan Yan, Jiaqi Chen, Xujie Zhang, Shengkai Zhang et al.ICCV 2021 · 16 citations
- 3D Human Pose Estimation via Explicit Compositional Depth MapsHaiping Wu, Bin XiaoAAAI 2020 · 23 citations
- BodyMap: Learning Full-Body Dense Correspondence MapAnastasia Ianina, Nikolaos Sarafianos, Yuanlu Xu, Ignacio Rocco et al.CVPR 2022 · 15 citations
- DeepCap: Monocular Human Performance Capture Using Weak SupervisionMarc Habermann, Weipeng Xu, Michael Zollhöfer, Gerard Pons-Moll et al.CVPR 2020
- Moulding Humans: Non-Parametric 3D Human Shape Estimation From Single ImagesValentin Gabeur, Jean-Sébastien Franco, Xavier Martin, Cordelia Schmid et al.ICCV 2019 · 140 citations
