Photorealistic Monocular 3D Reconstruction of Humans Wearing Clothing
Thiemo Alldieck, Mihai Zanfir, Cristian Sminchisescu
Abstract
We present PHORHUM, a novel, end-to-end trainable, deep neural network methodology for photorealistic 3D human reconstruction given just a monocular RGB image. Our pixel-aligned method estimates detailed 3D geometry and, for the first time, the unshaded surface color together with the scene illumination. Observing that 3D supervision alone is not sufficient for high fidelity color reconstruction, we introduce patch-based rendering losses that enable reliable color reconstruction on visible parts of the human, and detailed and plausible color estimation for the non-visible parts. Moreover, our method specifically addresses methodological and practical limitations of prior work in terms of representing geometry, albedo, and illumination effects, in an end-to-end model where factors can be effectively disentangled. In extensive experiments, we demonstrate the versatility and robustness of our approach. Our state-of-the-art results validate the method qualitatively and for different metrics, for both geometric and color reconstruction.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 306bc0c5-4d76-4bd5-970c-fb65562d85faCited by top-tier papers71
- SHERF: Generalizable Human NeRF from a Single ImageShoukang Hu, Fangzhou Hong, Liang Pan, Haiyi Mei et al.ICCV 2023 · 115 citations
- HumanSplat: Generalizable Single-Image Human Gaussian Splatting with Structure PriorsPanwang Pan, Zhuo Su, Chenguo Lin, Zhen Fan et al.NeurIPS 2024 · 76 citations
- Learning Clothing and Pose Invariant 3D Shape Representation for Long-Term Person Re-IdentificationFeng Liu, Minchul Kim, ZiAng Gu, Anil Jain et al.ICCV 2023 · 69 citations
- Global-correlated 3D-decoupling Transformer for Clothed Avatar ReconstructionZechuan Zhang, Li Sun, Zongxin Yang, Ling Chen et al.NeurIPS 2023 · 67 citations
- FOF: Learning Fourier Occupancy Field for Monocular Real-time Human ReconstructionQiao Feng, Yebin Liu, Yu-Kun Lai, Jingyu Yang et al.NeurIPS 2022 · 52 citations
Builds on20
- Fourier Features Let Networks Learn High Frequency Functions in Low Dimensional DomainsMatthew Tancik, Pratul P. Srinivasan, Ben Mildenhall, Sara Fridovich-Keil et al.NeurIPS 2020 · 4,036 citations
- PIFu: Pixel-Aligned Implicit Function for High-Resolution Clothed Human DigitizationShunsuke Saito, Zeng Huang, Ryota Natsume, Shigeo Morishima et al.ICCV 2019 · 1,411 citations
- Multiview Neural Surface Reconstruction by Disentangling Geometry and AppearanceLior Yariv, Yoni Kasten, Dror Moran, Meirav Galun et al.NeurIPS 2020 · 1,010 citations
- Implicit Geometric Regularization for Learning ShapesAmos Gropp, Lior Yariv, Niv Haim, Matan Atzmon et al.ICML 2020 · 1,001 citations
- Multi-Garment Net: Learning to Dress 3D People From ImagesBharat Lal Bhatnagar, Garvita Tiwari, Christian Theobalt, Gerard Pons-MollICCV 2019 · 447 citations
Related papers
- ARCH: Animatable Reconstruction of Clothed HumansZeng Huang, Yuanlu Xu, Christoph Lassner, Hao Li et al.CVPR 2020
- Monocular Real-Time Full Body Capture With Inter-Part CorrelationsYuxiao Zhou, Marc Habermann, Ikhsanul Habibie, Ayush Tewari et al.CVPR 2021
- IntrinsicAvatar: Physically Based Inverse Rendering of Dynamic Humans from Monocular Videos via Explicit Ray TracingShaofei Wang, Bozidar Antic, Andreas Geiger, Siyu TangCVPR 2024 · 13 citations
- Lighthouse: Predicting Lighting Volumes for Spatially-Coherent IlluminationPratul P. Srinivasan, Ben Mildenhall, Matthew Tancik, Jonathan T. Barron et al.CVPR 2020
- Uncalibrated Neural Inverse Rendering for Photometric Stereo of General SurfacesBerk Kaya, Suryansh Kumar, Carlos E. P. de Oliveira, Vittorio Ferrari et al.CVPR 2021
