Preface: A Data-driven Volumetric Prior for Few-shot Ultra High-resolution Face Synthesis
Marcel C. Bühler, Kripasindhu Sarkar, Tanmay Shah, Gengyan Li, Daoye Wang, Leonhard Helminger, Sergio Orts-Escolano, Dmitry Lagun, Otmar Hilliges, Thabo Beeler, Abhimitra Meka
Abstract
NeRFs have enabled highly realistic synthesis of human faces including complex appearance and reflectance effects of hair and skin. These methods typically require a large number of multi-view input images, making the process hardware intensive and cumbersome, limiting applicability to unconstrained settings. We propose a novel volumetric human face prior that enables the synthesis of ultra high-resolution novel views of subjects that are not part of the prior’s training distribution. This prior model consists of an identity-conditioned NeRF, trained on a dataset of low-resolution multi-view images of diverse humans with known camera calibration. A simple sparse landmark-based 3D alignment of the training dataset allows our model to learn a smooth latent space of geometry and appearance despite a limited number of training identities. A high-quality volumetric representation of a novel subject can be obtained by model fitting to 2 or 3 camera views of arbitrary resolution. Importantly, our method requires as few as two views of casually captured images as input at inference time.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b11b4cea-c3c9-4b6b-874a-b9b4b58ce11cCited by top-tier papers19
- Lite2Relight: 3D-aware Single Image Portrait RelightingPramod Rao, Gereon Fox, Abhimitra Meka, Mallikarjun B. R. et al.SIGGRAPH 2024 · 14 citations
- VRMM: A Volumetric Relightable Morphable Head ModelHaotian Yang, Mingwu Zheng, Chongyang Ma, Yu-Kun Lai et al.SIGGRAPH 2024 · 8 citations
- OHTA: One-shot Hand Avatar via Data-driven Implicit PriorsXiaozheng Zheng, Chao Wen, Zhuo Su, Zeran Xu et al.CVPR 2024 · 7 citations
- Learning Interaction-aware 3D Gaussian Splatting for One-shot Hand AvatarsXuan Huang, Hanhui Li, Wanquan Liu, Xiaodan Liang et al.NeurIPS 2024 · 6 citations
- FaceLift: Learning Generalizable Single Image 3D Face Reconstruction From Synthetic HeadsWeijie Lyu, Yi Zhou, Ming-Hsuan Yang, Zhixin ShuICCV 2025 · 5 citations
Builds on38
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Mip-NeRF: A Multiscale Representation for Anti-Aliasing Neural Radiance FieldsJonathan T. Barron, Ben Mildenhall, Matthew Tancik, Peter Hedman et al.ICCV 2021 · 2,700 citations
- NeuS: Learning Neural Implicit Surfaces by Volume Rendering for Multi-view ReconstructionPeng Wang, Lingjie Liu, Yuan Liu, Christian Theobalt et al.NeurIPS 2021 · 2,500 citations
- Mip-NeRF 360: Unbounded Anti-Aliased Neural Radiance FieldsJonathan T. Barron, Ben Mildenhall, Dor Verbin, Pratul P. Srinivasan et al.CVPR 2022 · 1,603 citations
- Volume Rendering of Neural Implicit SurfacesLior Yariv, Jiatao Gu, Yoni Kasten, Yaron LipmanNeurIPS 2021 · 1,421 citations
Related papers
- ViP-NeRF: Visibility Prior for Sparse Input Neural Radiance FieldsNagabhushan Somraj, Rajiv SoundararajanSIGGRAPH 2023 · 38 citations
- GM-NeRF: Learning Generalizable Model-Based Neural Radiance Fields from Multi-View ImagesJianchuan Chen, Wentao Yi, Liqian Ma, Xu Jia et al.CVPR 2023
- HumanNeRF: Efficiently Generated Human Radiance Field from Sparse InputsFuqiang Zhao, Wei Yang, Jiakai Zhang, Pei Lin et al.CVPR 2022 · 109 citations
- RegNeRF: Regularizing Neural Radiance Fields for View Synthesis from Sparse InputsMichael Niemeyer, Jonathan T. Barron, Ben Mildenhall, Mehdi S. M. Sajjadi et al.CVPR 2022 · 513 citations
- Pixel-Aligned Volumetric AvatarsAmit Raj, Michael Zollhöfer, Tomas Simon, Jason M. Saragih et al.CVPR 2021
