3D Face Tracking from 2D Video through Iterative Dense UV to Image Flow
Felix Taubner, Prashant Raina, Mathieu Tuli, Eu Wern Teh, Chul Lee, Jinmiao Huang
摘要
When working with 3D facial data, improving fidelity and avoiding the uncanny valley effect is critically dependent on accurate 3D facial performance capture. Because such methods are expensive and due to the widespread availability of 2D videos, recent methods have focused on how to perform monocular 3D face tracking. However, these methods often fall short in capturing precise facial movements due to limitations in their network architecture, training, and evaluation processes. Addressing these challenges, we propose a novel face tracker, FlowFace, that in-troduces an innovative 2D alignment network for dense pervertex alignment. Unlike prior work, FlowFace is trained on high-quality 3D scan annotations rather than weak supervision or synthetic data. Our 3D model fitting module Jointly fits a 3D face model from one or many observations, integrating existing neutral shape priors for enhanced identity and expression disentanglement and per-vertex de-formations for detailed facial feature reconstruction. Additionally, we propose a novel metric and benchmark for assessing tracking accuracy. Our method exhibits superior performance on both custom and publicly available bench-marks. We further validate the effectiveness of our tracker by generating high-quality 3D data from 2D videos, which leads to performance gains on downstream tasks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- Pixel3DMM: Versatile Screen-Space Priors for Single-Image 3D Face ReconstructionSimon Giebenhain, Tobias Kirschstein, Martin Rünz, Lourdes Agapito 等ICLR 2026 · 被引用 24 次
- Registration-Free Learnable Multi-View Capture of Faces in Dense Semantic CorrespondencePanagiotis Paraskevas Filntisis, George Retsinas, Radek Danecek, Vanessa Sklyarova 等CVPR 2026 · 被引用 3 次
- SHeaP: Self-Supervised Head Geometry Predictor Learned via 2D GaussiansLiam Schoneveld, Zhe Chen, Davide Davoli, Jiapeng Tang 等ICCV 2025 · 被引用 3 次
- PhysHead: Simulation-Ready Gaussian Head AvatarsBerna Kabadayi, Vanessa Sklyarova, Wojciech Zielonka, Justus Thies 等CVPR 2026 · 被引用 1 次
- Feed-forward Gaussian Registration for Head Avatar Creation and EditingMalte Prinzler, Paulo F. U. Gotardo, Siyu Tang, Timo BolkartCVPR 2026 · 被引用 1 次
它引用的顶会 Paper16
- SegFormer: Simple and Efficient Design for Semantic Segmentation with TransformersEnze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar 等NeurIPS 2021 · 被引用 9,661 次
- Learning an animatable detailed 3D face model from in-the-wild imagesYao Feng, Haiwen Feng, Michael J. Black, Timo BolkartSIGGRAPH 2021 · 被引用 662 次
- FaceFormer: Speech-Driven 3D Facial Animation with TransformersYingruo Fan, Zhaojiang Lin, Jun Saito, Wenping Wang 等CVPR 2022 · 被引用 218 次
- EMOCA: Emotion Driven Monocular Face Capture and AnimationRadek Danecek, Michael J. Black, Timo BolkartCVPR 2022 · 被引用 180 次
- Neural Head Avatars from Monocular RGB VideosPhilip-William Grassal, Malte Prinzler, Titus Leistner, Carsten Rother 等CVPR 2022 · 被引用 173 次
相关 Paper
- Accurate 3D Face Reconstruction with Facial Component TokensTianke Zhang, Xuangeng Chu, Yunfei Liu, Lijian Lin 等ICCV 2023 · 被引用 38 次
- DeepFaceFlow: In-the-Wild Dense 3D Facial Motion EstimationMohammad Rami Koujan, Anastasios Roussos, Stefanos ZafeiriouCVPR 2020
- Face Video Deblurring Using 3D Facial PriorsWenqi Ren, Jiaolong Yang, Senyou Deng, David P. Wipf 等ICCV 2019 · 被引用 52 次
- DGTalker: Disentangled Generative Latent Space Learning for Audio-Driven Gaussian Talking HeadsXiaoxi Liang, Yanbo Fan, Qiya Yang, Xuan Wang 等ICCV 2025 · 被引用 2 次
- JR2Net: Joint Monocular 3D Face Reconstruction and ReenactmentJiaxiang Shang, Yu Zeng, Xin Qiao, Xin Wang 等AAAI 2023 · 被引用 4 次
