Geometry-aware Two-scale PIFu Representation for Human Reconstruction
Zheng Dong, Ke Xu, Ziheng Duan, Hujun Bao, Weiwei Xu, Rynson W. H. Lau
Abstract
Although PIFu-based 3D human reconstruction methods are popular, the quality of recovered details is still unsatisfactory. In a sparse (e.g., 3 RGBD sensors) capture setting, the depth noise is typically amplified in the PIFu representation, resulting in flat facial surfaces and geometry-fallible bodies. In this paper, we propose a novel geometry-aware two-scale PIFu for 3D human reconstruction from sparse, noisy inputs. Our key idea is to exploit the complementary properties of depth denoising and 3D reconstruction, for learning a two-scale PIFu representation to reconstruct high-frequency facial details and consistent bodies separately. To this end, we first formulate depth denoising and 3D reconstruction as a multi-task learning problem. The depth denoising process enriches the local geometry information of the reconstruction features, while the reconstruction process enhances depth denoising with global topology information. We then propose to learn the two-scale PIFu representation using two MLPs based on the denoised depth and geometry-aware features. Extensive experiments demonstrate the effectiveness of our approach in reconstructing facial details and bodies of different poses and its superiority over state-of-the-art methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d1aa91e4-56f6-4aa7-926f-6eb76bc3fa39Cited by top-tier papers3
- ANIM: Accurate Neural Implicit Model for Human Reconstruction from a Single RGB-D ImageMarco Pesavento, Yuanlu Xu, Nikolaos Sarafianos, Robert Maier et al.CVPR 2024 · 11 citations
- Effective Video Mirror Detection with Inconsistent Motion CuesAlex Warren, Ke Xu, Jiaying Lin, Gary K. L. Tam et al.CVPR 2024 · 8 citations
- Video Mirror Detection with the Motion-in-Depth CueAlex Warren, Ke Xu, Xin Tian, Gary K. L. Tam et al.AAAI 2026
Builds on26
- PIFu: Pixel-Aligned Implicit Function for High-Resolution Clothed Human DigitizationShunsuke Saito, Zeng Huang, Ryota Natsume, Shigeo Morishima et al.ICCV 2019 · 1,411 citations
- DeepHuman: 3D Human Reconstruction From a Single ImageZerong Zheng, Tao Yu, Yixuan Wei, Qionghai Dai et al.ICCV 2019 · 367 citations
- Tex2Shape: Detailed Full Human Body Geometry From a Single ImageThiemo Alldieck, Gerard Pons-Moll, Christian Theobalt, Marcus A. MagnorICCV 2019 · 343 citations
- ARCH++: Animation-Ready Clothed Human Reconstruction RevisitedTong He, Yuanlu Xu, Shunsuke Saito, Stefano Soatto et al.ICCV 2021 · 233 citations
- Geo-PIFu: Geometry and Pixel Aligned Implicit Functions for Single-view Human ReconstructionTong He, John P. Collomosse, Hailin Jin, Stefano SoattoNeurIPS 2020 · 203 citations
Related papers
- StereoPIFu: Depth Aware Clothed Human Digitization via Stereo VisionYang Hong, Juyong Zhang, Boyi Jiang, Yudong Guo et al.CVPR 2021
- Robust 3D Self-Portraits in SecondsZhe Li, Tao Yu, Chuanyu Pan, Zerong Zheng et al.CVPR 2020
- DeepMultiCap: Performance Capture of Multiple Characters Using Sparse Multiview CamerasYang Zheng, Ruizhi Shao, Yuxiang Zhang, Tao Yu et al.ICCV 2021 · 112 citations
- SFDM: Robust Decomposition of Geometry and Reflectance for Realistic Face Rendering from Sparse-view ImagesDaisheng Jin, Jiangbei Hu, Baixin Xu, Yuxin Dai et al.CVPR 2025
- Deep Unsupervised 3D SfM Face Reconstruction Based on Massive Landmark Bundle AdjustmentYuxing Wang, Yawen Lu, Zhihua Xie, Guoyu LuACM MM 2021 · 15 citations
