High-fidelity Facial Avatar Reconstruction from Monocular Video with Generative Priors
Yunpeng Bai, Yanbo Fan, Xuan Wang, Yong Zhang, Jingxiang Sun, Chun Yuan, Ying Shan
Abstract
Figure 1 . Visualizations of 3DMM and audio-driven face reenactment of our proposed method and NerFACE [11] and DFRF [41] . The leftmost column is the ground truth image. For each method, the left plot is the rendered image with the same view of the ground truth image, and the two right plots are novel view syntheses. By utilizing the high-quality 3D-aware generative prior, our method significantly boosts the performance of face reenactment and novel view synthesis. We highlight some areas with red rectangles for better comparisons.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers17
- Generalizable and Animatable Gaussian Head AvatarXuangeng Chu, Tatsuya HaradaNeurIPS 2024 · 115 citations
- Fully Explicit Dynamic Gaussian SplattingJunoh Lee, Changyeon Won, Hyunjun Jung, Inhwan Bae et al.NeurIPS 2024 · 92 citations
- LAM: Large Avatar Model for One-shot Animatable Gaussian HeadYisheng He, Xiaodong Gu, Xiaodan Ye, Chao Xu et al.SIGGRAPH 2025 · 14 citations
- UniLS: End-to-End Audio-Driven Avatars for Unified Listening and SpeakingXuangeng Chu, Ruicong Liu, Yifei Huang, Yun Liu et al.CVPR 2026 · 12 citations
- Generalizable One-shot 3D Neural Head AvatarXueting Li, Shalini De Mello, Sifei Liu, Koki Nagano et al.NeurIPS 2023 · 12 citations
Builds on27
- Nerfies: Deformable Neural Radiance FieldsKeunhong Park, Utkarsh Sinha, Jonathan T. Barron, Sofien Bouaziz et al.ICCV 2021 · 1,442 citations
- Efficient Geometry-aware 3D Generative Adversarial NetworksEric R. Chan, Connor Z. Lin, Matthew A. Chan, Koki Nagano et al.CVPR 2022 · 984 citations
- A Lip Sync Expert Is All You Need for Speech to Lip Generation In the WildK. R. Prajwal, Rudrabha Mukhopadhyay, Vinay P. Namboodiri, C. V. JawaharACM MM 2020 · 869 citations
- FSGAN: Subject Agnostic Face Swapping and ReenactmentYuval Nirkin, Yosi Keller, Tal HassnerICCV 2019 · 710 citations
- StyleNeRF: A Style-based 3D Aware Generator for High-resolution Image SynthesisJiatao Gu, Lingjie Liu, Peng Wang, Christian TheobaltICLR 2022 · 622 citations
Related papers
- VOODOO 3D: Volumetric Portrait Disentanglement for One-Shot 3D Head ReenactmentPhong Tran, Egor Zakharov, Long-Nhat Ho, Anh Tuan Tran et al.CVPR 2024 · 15 citations
- GPAvatar: High-fidelity Head Avatars by Learning Efficient Gaussian ProjectionsWei-Qi Feng, Dong Han, Ze-Kang Zhou, Shunkai Li et al.CVPR 2025
- ReDA: Reinforced Differentiable Attribute for 3D Face ReconstructionWenbin Zhu, HsiangTao Wu, Zeyu Chen, Noranart Vesdapunt et al.CVPR 2020
- Learning Dense Correspondence for NeRF-Based Face ReenactmentSonglin Yang, Wei Wang, Yushi Lan, Xiangyu Fan et al.AAAI 2024 · 17 citations
- PIRenderer: Controllable Portrait Image Generation via Semantic Neural RenderingYurui Ren, Ge Li, Yuanqi Chen, Thomas H. Li et al.ICCV 2021 · 284 citations
