Implicit Neural Head Synthesis via Controllable Local Deformation Fields
Chuhan Chen, Matthew O'Toole, Gaurav Bharaj, Pablo Garrido
Abstract
High-quality reconstruction of controllable 3D head avatars from 2D videos is highly desirable for virtual human applications in movies, games, and telepresence. Neural implicit fields provide a powerful representation to model 3D head avatars with personalized shape, expressions, and facial parts, e.g., hair and mouth interior, that go beyond the linear 3D morphable model (3DMM). However, existing methods do not model faces with fine-scale facial features, or local control of facial parts that extrapolate asymmetric expressions from monocular videos. Further, most condition only on 3DMM parameters with poor(er) locality, and resolve local features with a global neural field. We build on part-based implicit shape models that decompose a global deformation field into local ones. Our novel formulation models multiple implicit deformation fields with local semantic rig-like control via 3DMM-based parameters, and representative facial landmarks. Further, we propose a local control loss and attention mask mechanism that promote sparsity of each learned deformation field. Our formulation renders sharper locally controllable nonlinear deformations than previous implicit monocular approaches, especially mouth interior, asymmetric expressions, and facial details. Project page: https://imaging.cs.cmu.edu/local deformation fields/
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 7f95e0a0-e662-4be6-a414-003cd677e3eaCited by top-tier papers5
- VOODOO 3D: Volumetric Portrait Disentanglement for One-Shot 3D Head ReenactmentPhong Tran, Egor Zakharov, Long-Nhat Ho, Anh Tuan Tran et al.CVPR 2024 · 15 citations
- Streamlined Facial Data Collection Based on Utterance and Emotional Data for Human-to-Avatar ReconstructionSeoyoung Kang, Seokhwan Yang, Hail Song, Boram Yoon et al.IEEE VR 2026
- 3D Gaussian Head Avatars with Expressive Dynamic Appearances by Compact Tensorial RepresentationsYating Wang, Xuan Wang, Ran Yi, Yanbo Fan et al.CVPR 2025
- MeGA: Hybrid Mesh-Gaussian Head Avatar for High-Fidelity Rendering and Head EditingCong Wang, Di Kang, Heyi Sun, Shen-Han Qian et al.CVPR 2025
- Efficient 3D Implicit Head Avatar With Mesh-Anchored Hash Table BlendshapesZiqian Bai, Feitong Tan, Sean Fanello, Rohit Pandey et al.CVPR 2024
Builds on34
- Fourier Features Let Networks Learn High Frequency Functions in Low Dimensional DomainsMatthew Tancik, Pratul P. Srinivasan, Ben Mildenhall, Sara Fridovich-Keil et al.NeurIPS 2020 · 4,036 citations
- Efficient Geometry-aware 3D Generative Adversarial NetworksEric R. Chan, Connor Z. Lin, Matthew A. Chan, Koki Nagano et al.CVPR 2022 · 984 citations
- Learning an animatable detailed 3D face model from in-the-wild imagesYao Feng, Haiwen Feng, Michael J. Black, Timo BolkartSIGGRAPH 2021 · 662 citations
- StyleNeRF: A Style-based 3D Aware Generator for High-resolution Image SynthesisJiatao Gu, Lingjie Liu, Peng Wang, Christian TheobaltICLR 2022 · 622 citations
- Non-Rigid Neural Radiance Fields: Reconstruction and Novel View Synthesis of a Dynamic Scene From Monocular VideoEdgar Tretschk, Ayush Tewari, Vladislav Golyanik, Michael Zollhöfer et al.ICCV 2021 · 617 citations
Related papers
- I M Avatar: Implicit Morphable Head Avatars from VideosYufeng Zheng, Victoria Fernández Abrevaya, Marcel C. Bühler, Xu Chen et al.CVPR 2022 · 169 citations
- MonoGaussianAvatar: Monocular Gaussian Point-based Head AvatarYufan Chen, Lizhen Wang, Qijing Li, Hongjiang Xiao et al.SIGGRAPH 2024 · 85 citations
- RigNeRF: Fully Controllable Neural 3D PortraitsShahRukh Athar, Zexiang Xu, Kalyan Sunkavalli, Eli Shechtman et al.CVPR 2022 · 117 citations
- PointAvatar: Deformable Point-Based Head Avatars from VideosYufeng Zheng, Wang Yifan, Gordon Wetzstein, Michael J. Black et al.CVPR 2023
- ImFace: A Nonlinear 3D Morphable Face Model with Implicit Neural RepresentationsMingwu Zheng, Hongyu Yang, Di Huang, Liming ChenCVPR 2022 · 60 citations
