Deep 3D Portrait From a Single Image
Sicheng Xu, Jiaolong Yang, Dong Chen, Fang Wen, Yu Deng, Yunde Jia, Xin Tong
Abstract
In this paper, we present a learning-based approach for recovering the 3D geometry of human head from a single portrait image. Our method is learned in an unsupervised manner without any ground-truth 3D data. We represent the head geometry with a parametric 3D face model together with a depth map for other head regions including hair and ear. A two-step geometry learning scheme is proposed to learn 3D head reconstruction from in-the-wild face images, where we first learn face shape on single images using selfreconstruction and then learn hair and ear geometry using pairs of images in a stereo-matching fashion. The second step is based on the output of the first to not only improve the accuracy but also ensure the consistency of overall head geometry. We evaluate the accuracy of our method both in 3D and with pose manipulation tasks on 2D images. We alter pose based on the recovered geometry and apply a refinement network trained with adversarial learning to ameliorate the reprojected images and translate them to the real image domain. Extensive evaluations and comparison with previous methods show that our new method can produce high-fidelity 3D head geometry and head pose manipulation results. * This work was done when S. Xu was an intern at MSRA. tates substantial image content regeneration in the head region and beyond. Promising results have been shown for face rotation [58, 3, 27] with generative adversarial nets (GAN), but generating the whole head region with new poses is still far from being solved. One reason could be implicitly learning the complex 3D geometry of a large variety of hair styles and interpret them onto 2D pixel grid is still prohibitively challenging for GANs.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 320de722-e6f1-4b1c-94b5-15f38cd28648Cited by top-tier papers23
- Generalizable and Animatable Gaussian Head AvatarXuangeng Chu, Tatsuya HaradaNeurIPS 2024 · 115 citations
- AniFaceGAN: Animatable 3D-Aware Face Image Generation for Video AvatarsYue Wu, Yu Deng, Jiaolong Yang, Fangyun Wei et al.NeurIPS 2022 · 77 citations
- VirtualCube: An Immersive 3D Video Communication SystemYizhong Zhang, Jiaolong Yang, Zhen Liu, Ruicheng Wang et al.IEEE VR 2022 · 66 citations
- FaceController: Controllable Attribute Editing for Face in the WildZhiliang Xu, Xiyu Yu, Zhibin Hong, Zhen Zhu et al.AAAI 2021 · 49 citations
- SOGAN: 3D-Aware Shadow and Occlusion Robust GAN for Makeup TransferYueming Lyu, Jing Dong, Bo Peng, Wei Wang et al.ACM MM 2021 · 36 citations
Builds on3
- FSGAN: Subject Agnostic Face Swapping and ReenactmentYuval Nirkin, Yosi Keller, Tal HassnerICCV 2019 · 710 citations
- Disentangled and Controllable Face Image Generation via 3D Imitative-Contrastive LearningYu Deng, Jiaolong Yang, Dong Chen, Fang Wen et al.CVPR 2020
- Interpreting the Latent Space of GANs for Semantic Face EditingYujun Shen, Jinjin Gu, Xiaoou Tang, Bolei ZhouCVPR 2020
Related papers
- Generalizable One-shot 3D Neural Head AvatarXueting Li, Shalini De Mello, Sifei Liu, Koki Nagano et al.NeurIPS 2023 · 12 citations
- Depth-Aware Generative Adversarial Network for Talking Head Video GenerationFa-Ting Hong, Longhao Zhang, Li Shen, Dan XuCVPR 2022 · 168 citations
- Learning to Restore 3D Face from In-the-Wild Degraded ImagesZhenyu Zhang, Yanhao Ge, Ying Tai, Xiaoming Huang et al.CVPR 2022 · 4 citations
- Do 2D GANs Know 3D Shape? Unsupervised 3D Shape Reconstruction from 2D Image GANsXingang Pan, Bo Dai, Ziwei Liu, Chen Change Loy et al.ICLR 2021 · 119 citations
- Im2Haircut: Single-View Strand-Based Hair Reconstruction for Human AvatarsVanessa Sklyarova, Egor Zakharov, Malte Prinzler, Giorgio Becherini et al.ICCV 2025 · 3 citations
