EMOCA: Emotion Driven Monocular Face Capture and Animation
Radek Danecek, Michael J. Black, Timo Bolkart
摘要
As 3D facial avatars become more widely used for communication, it is critical that they faithfully convey emotion. Unfortunately, the best recent methods that regress parametric 3D face models from monocular images are unable to capture the full spectrum of facial expression, such as subtle or extreme emotions. We find the standard reconstruction metrics used for training (landmark reprojection error, photometric error, and face recognition loss) are insufficient to capture high-fidelity expressions. The result is facial geometries that do not match the emotional content of the input image. We address this with EMOCA (EMOtion Capture and Animation), by introducing a novel deep perceptual emotion consistency loss during training, which helps ensure that the reconstructed 3D expression matches the expression depicted in the input image. While EMOCA achieves 3D reconstruction errors that are on par with the current best methods, it significantly outperforms them in terms of the quality of the reconstructed expression and the perceived emotional content. We also directly regress levels of valence and arousal and classify basic expressions from the estimated 3D face parameters. On the task of in-the-wild emotion recognition, our purely geometric approach is on par with the best image-based methods, highlighting the value of 3D geometry in analyzing human behavior. The model and code are publicly available at https://emoca.is.tue.mpg.de.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper104
- Real3D-Portrait: One-shot Realistic 3D Talking Portrait SynthesisZhenhui Ye, Tianyun Zhong, Yi Ren, Jiaqi Yang 等ICLR 2024 · 被引用 105 次
- AvatarReX: Real-time Expressive Full-body AvatarsZerong Zheng, Xiaochen Zhao, Hongwen Zhang, Boning Liu 等SIGGRAPH 2023 · 被引用 80 次
- GPAvatar: Generalizable and Precise Head Avatar from Image(s)Xuangeng Chu, Yu Li, Ailing Zeng, Tianyu Yang 等ICLR 2024 · 被引用 63 次
- Can Language Models Learn to Listen?Evonne Ng, Sanjay Subramanian, Dan Klein, Angjoo Kanazawa 等ICCV 2023 · 被引用 44 次
- HiFace: High-Fidelity 3D Face Reconstruction by Learning Static and Dynamic DetailsZenghao Chai, Tianke Zhang, Tianyu He, Xu Tan 等ICCV 2023 · 被引用 33 次
它引用的顶会 Paper8
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- Learning an animatable detailed 3D face model from in-the-wild imagesYao Feng, Haiwen Feng, Michael J. Black, Timo BolkartSIGGRAPH 2021 · 被引用 662 次
- Racial Faces in the Wild: Reducing Racial Bias by Information Maximization Adaptation NetworkMei Wang, Weihong Deng, Jiani Hu, Xunqiang Tao 等ICCV 2019 · 被引用 379 次
- DF2Net: A Dense-Fine-Finer Network for Detailed 3D Face ReconstructionXiaoxing Zeng, Xiaojiang Peng, Yu QiaoICCV 2019 · 被引用 85 次
- Unsupervised Learning of Probably Symmetric Deformable 3D Objects From Images in the WildShangzhe Wu, Christian Rupprecht, Andrea VedaldiCVPR 2020
相关 Paper
- DeepFaceFlow: In-the-Wild Dense 3D Facial Motion EstimationMohammad Rami Koujan, Anastasios Roussos, Stefanos ZafeiriouCVPR 2020
- Neural Emotion Director: Speech-preserving semantic control of facial expressions in "in-the-wild" videosFoivos Paraperas Papantoniou, Panagiotis Paraskevas Filntisis, Petros Maragos, Anastasios RoussosCVPR 2022 · 被引用 32 次
- 3D Facial Expressions through Analysis-by-Neural-SynthesisGeorge Retsinas, Panagiotis Paraskevas Filntisis, Radek Danecek, Victoria Fernández Abrevaya 等CVPR 2024 · 被引用 26 次
- Uncertainty-Aware Semi-Supervised Learning of 3D Face Rigging from Single ImageYong Zhao, Haifeng Chen, Hichem Sahli, Ke Lu 等ACM MM 2022 · 被引用 2 次
- Towards Accurate Facial Motion Retargeting with Identity-Consistent and Expression-Exclusive ConstraintsLangyuan Mo, Haokun Li, Chaoyang Zou, Yubing Zhang 等AAAI 2022 · 被引用 9 次
