HeadGAN: One-shot Neural Head Synthesis and Editing
Michail Christos Doukas, Stefanos Zafeiriou, Viktoriia Sharmanska
Abstract
Recent attempts to solve the problem of head reenactment using a single reference image have shown promising results. However, most of them either perform poorly in terms of photo-realism, or fail to meet the identity preservation problem, or do not fully transfer the driving pose and expression. We propose HeadGAN, a novel system that conditions synthesis on 3D face representations, which can be extracted from any driving video and adapted to the facial geometry of any reference image, disentangling identity from expression. We further improve mouth movements, by utilising audio features as a complementary input. The 3D face representation enables HeadGAN to be further used as an efficient method for compression and reconstruction and a tool for expression and pose editing.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers55
- Thin-Plate Spline Motion Model for Image AnimationJian Zhao, Hui ZhangCVPR 2022 · 196 citations
- Expressive Talking Head Generation with Granular Audio-Visual ControlBorong Liang, Yan Pan, Zhizhi Guo, Hang Zhou et al.CVPR 2022 · 114 citations
- DINet: Deformation Inpainting Network for Realistic Face Visually Dubbing on High Resolution VideoZhimeng Zhang, Zhipeng Hu, Wenjin Deng, Changjie Fan et al.AAAI 2023 · 106 citations
- MegaPortraits: One-shot Megapixel Neural Head AvatarsNikita Drobyshev, Jenya Chelishev, Taras Khakhulin, Aleksei Ivakhnenko et al.ACM MM 2022 · 86 citations
- AvatarMAV: Fast 3D Head Avatar Reconstruction Using Motion-Aware Neural VoxelsYuelang Xu, Lizhen Wang, Xiaochen Zhao, Hongwen Zhang et al.SIGGRAPH 2023 · 69 citations
Builds on7
- Few-Shot Adversarial Learning of Realistic Neural Talking Head ModelsEgor Zakharov, Aliaksandra Shysheya, Egor Burkov, Victor S. LempitskyICCV 2019 · 687 citations
- MarioNETte: Few-Shot Face Reenactment Preserving Identity of Unseen TargetsSungjoo Ha, Martin Kersner, Beomsu Kim, Seokjun Seo et al.AAAI 2020 · 184 citations
- Encoding in Style: A StyleGAN Encoder for Image-to-Image TranslationElad Richardson, Yuval Alaluf, Or Patashnik, Yotam Nitzan et al.CVPR 2021
- RetinaFace: Single-Shot Multi-Level Face Localisation in the WildJiankang Deng, Jia Guo, Evangelos Ververas, Irene Kotsia et al.CVPR 2020
- DeepFaceFlow: In-the-Wild Dense 3D Facial Motion EstimationMohammad Rami Koujan, Anastasios Roussos, Stefanos ZafeiriouCVPR 2020
Related papers
- Enhancing Identity-Deformation Disentanglement in StyleGAN for One-Shot Face Video Re-EnactmentQing Chang, Yao-Xiang Ding, Kun ZhouAAAI 2025 · 3 citations
- Neural Head Reenactment with Latent Pose DescriptorsEgor Burkov, Igor Pasechnik, Artur Grigorev, Victor S. LempitskyCVPR 2020
- Talking Face Generation with Expression-Tailored Generative Adversarial NetworkDan Zeng, Han Liu, Hui Lin, Shiming GeACM MM 2020 · 30 citations
- Coherent 3D Portrait Video Reconstruction via Triplane FusionShengze Wang, Xueting Li, Chao Liu, Matthew A. Chan et al.CVPR 2025
- Mesh Guided One-shot Face Reenactment Using Graph Convolutional NetworksGuangming Yao, Yi Yuan, Tianjia Shao, Kun ZhouACM MM 2020 · 42 citations
