Neural Emotion Director: Speech-preserving semantic control of facial expressions in "in-the-wild" videos
Foivos Paraperas Papantoniou, Panagiotis Paraskevas Filntisis, Petros Maragos, Anastasios Roussos
摘要
In this paper, we introduce a novel deep learning method for photo-realistic manipulation of the emotional state of actors in “in-the-wild” videos. The proposed method is based on a parametric 3D face representation of the actor in the input scene that offers a reliable disentanglement of the facial identity from the head pose and facial expressions. It then uses a novel deep domain translation framework that alters the facial expressions in a consistent and plausible manner, taking into account their dynamics. Finally, the altered facial expressions are used to photo-realistically manipulate the facial region in the input scene based on an especially-designed neural face renderer. To the best of our knowledge, our method is the first to be capable of controlling the actor's facial expressions by even using as a sole input the semantic labels of the manipulated emotions, while at the same time preserving the speech-related lip movements. We conduct extensive qualitative and quantitative evaluations and comparisons, which demonstrate the effectiveness of our approach and the especially promising results that we obtain. Our method opens a plethora of new possibilities for useful applications of neural rendering technologies, ranging from movie post-production and video games to photo-realistic affective avatars.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- EmoTalk: Speech-Driven Emotional Disentanglement for 3D Face AnimationZiqiao Peng, Haoyu Wu, Zhenbo Song, Hao Xu 等ICCV 2023 · 被引用 192 次
- 3D Facial Expressions through Analysis-by-Neural-SynthesisGeorge Retsinas, Panagiotis Paraskevas Filntisis, Radek Danecek, Victoria Fernández Abrevaya 等CVPR 2024 · 被引用 26 次
- A Unified and Interpretable Emotion Representation and Expression GenerationReni Paskaleva, Mykyta Holubakha, Andela Ilic, Saman Motamed 等CVPR 2024 · 被引用 5 次
- When Words Smile: Generating Diverse Emotional Facial Expressions from TextHaidong Xu, Meishan Zhang, Hao Ju, Zhedong Zheng 等EMNLP 2025 · 被引用 2 次
- Learning to Dub Movies via Hierarchical Prosody ModelsGaoxiang Cong, Liang Li, Yuankai Qi, Zheng-Jun Zha 等CVPR 2023
它引用的顶会 Paper6
- FSGAN: Subject Agnostic Face Swapping and ReenactmentYuval Nirkin, Yosi Keller, Tal HassnerICCV 2019 · 被引用 710 次
- Few-Shot Adversarial Learning of Realistic Neural Talking Head ModelsEgor Zakharov, Aliaksandra Shysheya, Egor Burkov, Victor S. LempitskyICCV 2019 · 被引用 687 次
- Learning an animatable detailed 3D face model from in-the-wild imagesYao Feng, Haiwen Feng, Michael J. Black, Timo BolkartSIGGRAPH 2021 · 被引用 662 次
- Audio-Driven Emotional Video PortraitsXinya Ji, Hang Zhou, Kaisiyuan Wang, Wayne Wu 等CVPR 2021
- GANmut: Learning Interpretable Conditional Space for Gamut of EmotionsStefano d'Apolito, Danda Pani Paudel, Zhiwu Huang, Andrés Romero 等CVPR 2021
相关 Paper
- Neural Face Rigging for Animating and Retargeting Facial Meshes in the WildDafei Qin, Jun Saito, Noam Aigerman, Thibault Groueix 等SIGGRAPH 2023 · 被引用 26 次
- FG-EmoTalk: Talking Head Video Generation with Fine-Grained Controllable Facial ExpressionsZhaoxu Sun, Yuze Xuan, Fang Liu, Yang XiangAAAI 2024 · 被引用 13 次
- Uncertainty-Aware Semi-Supervised Learning of 3D Face Rigging from Single ImageYong Zhao, Haifeng Chen, Hichem Sahli, Ke Lu 等ACM MM 2022 · 被引用 2 次
- VariTex: Variational Neural Face TexturesMarcel C. Bühler, Abhimitra Meka, Gengyan Li, Thabo Beeler 等ICCV 2021 · 被引用 43 次
- TACR-Net: Editing on Deep Video and Voice PortraitsLuchuan Song, Bin Liu, Guojun Yin, Xiaoyi Dong 等ACM MM 2021 · 被引用 20 次
