Self-Supervised 3D Face Reconstruction via Conditional Estimation
Yandong Wen, Weiyang Liu, Bhiksha Raj, Rita Singh
Abstract
We present a conditional estimation (CEST) framework to learn 3D facial parameters from 2D single-view images by self-supervised training from videos. CEST is based on the process of analysis by synthesis, where the 3D facial parameters (shape, reflectance, viewpoint, and illumination) are estimated from the face image, and then recombined to reconstruct the 2D face image. In order to learn semantically meaningful 3D facial parameters without explicit access to their labels, CEST couples the estimation of different 3D facial parameters by taking their statistical dependency into account. Specifically, the estimation of any 3D facial parameter is not only conditioned on the given image, but also on the facial parameters that have already been derived. Moreover, the reflectance symmetry and consistency among the video frames are adopted to improve the disentanglement of facial parameters. Together with a novel strategy for incorporating the reflectance symmetry and consistency, CEST can be efficiently trained with in-the-wild video clips. Both qualitative and quantitative experiments demonstrate the effectiveness of CEST.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 971efd55-8482-41b6-9054-3a6742127b5bCited by top-tier papers11
- H3WB: Human3.6M 3D WholeBody Dataset and BenchmarkYue Zhu, Nermin Samet, David PicardICCV 2023 · 34 citations
- HiFace: High-Fidelity 3D Face Reconstruction by Learning Static and Dynamic DetailsZenghao Chai, Tianke Zhang, Tianyu He, Xu Tan et al.ICCV 2023 · 33 citations
- iVS-Net: Learning Human View Synthesis from Internet VideosJunting Dong, Qi Fang, Tianshuo Yang, Qing Shuai et al.ICCV 2023 · 9 citations
- Makeup Prior Models for 3D Facial Makeup Estimation and ApplicationsXingchao Yang, Takafumi Taketomi, Yuki Endo, Yoshihiro KanamoriCVPR 2024 · 7 citations
- Universal Facial Encoding of Codec Avatars from VR HeadsetsShaojie Bai, Te-Li Wang, Chenghui Li, Akshay Venkatesh et al.SIGGRAPH 2024 · 4 citations
Builds on2
Related papers
- Exploiting Self-Supervised and Semi-Supervised Learning for Facial Landmark Tracking with Unlabeled DataShi Yin, Shangfei Wang, Xiaoping Chen, Enhong ChenACM MM 2020 · 7 citations
- End-to-End 3D Face Reconstruction with Expressions and Specular Albedos from Single In-the-wild ImagesQixin Deng, Binh Huy Le, Aobo Jin, Zhigang DengACM MM 2022
- Deep Unsupervised 3D SfM Face Reconstruction Based on Massive Landmark Bundle AdjustmentYuxing Wang, Yawen Lu, Zhihua Xie, Guoyu LuACM MM 2021 · 15 citations
- DeepFaceFlow: In-the-Wild Dense 3D Facial Motion EstimationMohammad Rami Koujan, Anastasios Roussos, Stefanos ZafeiriouCVPR 2020
- LipSync3D: Data-Efficient Learning of Personalized 3D Talking Faces From Video Using Pose and Lighting NormalizationAvisek Lahiri, Vivek Kwatra, Christian Früh, John Lewis et al.CVPR 2021
