FaceLift: Learning Generalizable Single Image 3D Face Reconstruction From Synthetic Heads
Weijie Lyu, Yi Zhou, Ming-Hsuan Yang, Zhixin Shu
Abstract
We present FaceLift, a novel feed-forward approach for generalizable high-quality 360-degree 3D head reconstruction from a single image. Our pipeline first employs a multiview latent diffusion model to generate consistent side and back views from a single facial input, which then feed into a transformer-based reconstructor that produces a comprehensive 3D Gaussian splats representation. Previous methods for monocular 3D face reconstruction often lack full view coverage or view consistency due to insufficient multi-view supervision. We address this by creating a highquality synthetic head dataset that enables consistent supervision across viewpoints. To bridge the domain gap between synthetic training data and real-world images, we propose a simple yet effective technique that ensures the view generation process maintains fidelity to the input by learning to reconstruct the input image alongside the view generation. Despite being trained exclusively on synthetic data, our method demonstrates remarkable generalization to real-world images. Through extensive qualitative and quantitative evaluations, we show that FaceLift outperforms state-of-the-art 3D face reconstruction methods on identity preservation, detail recovery and rendering quality.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers9
- FlexAvatar: Learning Complete 3D Head Avatars with Partial SupervisionTobias Kirschstein, Simon Giebenhain, Matthias NießnerCVPR 2026 · 10 citations
- PercHead: Perceptual Head Model for Single-Image 3D Head Reconstruction & EditingAntonio Oroz, Matthias Nießner, Tobias KirschsteinCVPR 2026 · 6 citations
- VoluMe - Authentic 3D Video Calls from Live Gaussian Splat PredictionMartin de La Gorce, Charlie Hewitt, Tibor Takács, Robert Gerdisch et al.ICCV 2025 · 5 citations
- Feed-forward Gaussian Registration for Head Avatar Creation and EditingMalte Prinzler, Paulo F. U. Gotardo, Siyu Tang, Timo BolkartCVPR 2026 · 1 citation
- TokenLight: Precise Lighting Control in Images using Attribute TokensSumit Chaturvedi, Yannick Hold-Geoffroy, Mengwei Ren, Jingyuan Liu et al.CVPR 2026 · 1 citation
Builds on36
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 5,687 citations
- NeuS: Learning Neural Implicit Surfaces by Volume Rendering for Multi-view ReconstructionPeng Wang, Lingjie Liu, Yuan Liu, Christian Theobalt et al.NeurIPS 2021 · 2,500 citations
Related papers
- Human-3Diffusion: Realistic Avatar Creation via Explicit 3D Consistent Diffusion ModelsYuxuan Xue, Xianghui Xie, Riccardo Marin, Gerard Pons-MollNeurIPS 2024 · 49 citations
- HumanSplat: Generalizable Single-Image Human Gaussian Splatting with Structure PriorsPanwang Pan, Zhuo Su, Chenguo Lin, Zhen Fan et al.NeurIPS 2024 · 76 citations
- High-Quality Full-Head 3D Avatar Generation from Any Single Portrait ImageYujie Gao, Chencheng Wang, Xianbing Sun, Jiahui Zhan et al.AAAI 2026
- Diff4Splat: Repurposing Video Diffusion Models for Dynamic Scene GenerationPanwang Pan, Chenguo Lin, Chenxin Li, Jingjing Zhao et al.CVPR 2026
- GAF: Gaussian Avatar Reconstruction from Monocular Videos via Multi-view DiffusionJiapeng Tang, Davide Davoli, Tobias Kirschstein, Liam Schoneveld et al.CVPR 2025
