Versatile Face Animator: Driving Arbitrary 3D Facial Avatar in RGBD Space
Haoyu Wang, Haozhe Wu, Junliang Xing, Jia Jia
Abstract
Creating realistic 3D facial animation is crucial for various applications in the movie production and gaming industry, especially with the burgeoning demand in the metaverse. However, prevalent methods such as blendshape-based approaches and facial rigging techniques are time-consuming, labor-intensive, and lack standardized configurations, making facial animation production challenging and costly. In this paper, we propose a novel self-supervised framework, Versatile Face Animator, which combines facial motion capture with motion retargeting in an end-to-end manner, eliminating the need for blendshapes or rigs. Our method has the following two main characteristics: 1) we propose an RGBD animation module to learn facial motion from raw RGBD videos by hierarchical motion dictionaries and animate RGBD images rendered from 3D facial mesh coarse-to-fine, enabling facial animation on arbitrary 3D characters regardless of their topology, textures, blendshapes, and rigs; and 2) we introduce a mesh retarget module to utilize RGBD animation to create 3D facial animation by manipulating facial mesh with controller transformations, which are estimated from dense optical flow fields and blended together with geodesic-distance-based weights. Comprehensive experiments demonstrate the effectiveness of our proposed framework in generating impressive 3D facial animation results, highlighting its potential as a promising solution for the cost-effective and efficient production of facial animation in the metaverse.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2ccae9ef-151d-4e9c-86c1-bd50903b0eddCited by top-tier papers3
- MMHead: Towards Fine-grained Multi-modal 3D Facial AnimationSijing Wu, Yunhao Li, Yichao Yan, Huiyu Duan et al.ACM MM 2024 · 17 citations
- VividListener: Expressive and Controllable Listener Dynamics Modeling for Multi-Modal Responsive InteractionShiying Li, Xingqun Qi, Bingkun Yang, Weile Chen et al.AAAI 2026 · 2 citations
- HarmoniVox: Painting Voices to Match the Avatar's SoulSongtao Zhou, Xiaoyu Qin, Yixuan Zhou, Qixin Wang et al.ACM MM 2025
Builds on14
- Few-Shot Adversarial Learning of Realistic Neural Talking Head ModelsEgor Zakharov, Aliaksandra Shysheya, Egor Burkov, Victor S. LempitskyICCV 2019 · 687 citations
- Learning an animatable detailed 3D face model from in-the-wild imagesYao Feng, Haiwen Feng, Michael J. Black, Timo BolkartSIGGRAPH 2021 · 662 citations
- Latent Image Animator: Learning to Animate Images via Latent Space NavigationYaohui Wang, Di Yang, François Brémond, Antitza DantchevaICLR 2022 · 219 citations
- Depth-Aware Generative Adversarial Network for Talking Head Video GenerationFa-Ting Hong, Longhao Zhang, Li Shen, Dan XuCVPR 2022 · 168 citations
- Fast and deep facial deformationsStephen W. Bailey, Dalton Omens, Paul C. DiLorenzo, James F. O'BrienSIGGRAPH 2020 · 45 citations
Related papers
- VASA-Rig: Audio-Driven 3D Facial Animation with 'Live' Mood Dynamics in Virtual RealityYe Pan, Chang Liu, Sicheng Xu, Shuai Tan et al.IEEE VR 2025 · 5 citations
- Learning a Generalized Physical Face Model From DataLingchen Yang, Gaspard Zoss, Prashanth Chandran, Markus Gross et al.SIGGRAPH 2024 · 10 citations
- Bring Your Own Character: A Holistic Solution for Automatic Facial Animation Generation of Customized CharactersZechen Bai, Peng Chen, Xiaolan Peng, Lu Liu et al.IEEE VR 2024 · 8 citations
- Local anatomically-constrained facial performance retargetingPrashanth Chandran, Loïc Ciccone, Markus Gross, Derek BradleySIGGRAPH 2022 · 25 citations
- Neural Face Rigging for Animating and Retargeting Facial Meshes in the WildDafei Qin, Jun Saito, Noam Aigerman, Thibault Groueix et al.SIGGRAPH 2023 · 26 citations
