Controllable 3D Face Synthesis with Conditional Generative Occupancy Fields
Keqiang Sun, Shangzhe Wu, Zhaoyang Huang, Ning Zhang, Quan Wang, Hongsheng Li
摘要
Capitalizing on the recent advances in image generation models, existing controllable face image synthesis methods are able to generate high-fidelity images with some levels of controllability, e.g., controlling the shapes, expressions, textures, and poses of the generated face images. However, these methods focus on 2D image generative models, which are prone to producing inconsistent face images under large expression and pose changes. In this paper, we propose a new NeRF-based conditional 3D face synthesis framework, which enables 3D controllability over the generated face images by imposing explicit 3D conditions from 3D face priors. At its core is a conditional Generative Occupancy Field (cGOF) that effectively enforces the shape of the generated face to commit to a given 3D Morphable Model (3DMM) mesh. To achieve accurate control over fine-grained 3D face shapes of the synthesized image, we additionally incorporate a 3D landmark loss as well as a volume warping loss into our synthesis algorithm. Experiments validate the effectiveness of the proposed method, which can generate high-fidelity face images and shows more precise 3D controllability than state-ofthe-art 2D-based controllable face synthesis methods. Find code and more demo at https://keqiangsun.github.io/projects/cgof . Recent success of Generative Adversarial Networks (GANs) [13] has led to tremendous progress in face image synthesis. State-of-the-art methods, such as StyleGAN [21, 22, 20] , are capable of generating photo-realistic face images. Apart from photo-realism, being able to control the appearance of the generated images is also key in many real-world applications, such as face animation, reenactment, and free-viewpoint rendering. Early works on controllable face synthesis rely on external attribute annotations to learn an attribute-guided face image generation model [27, 11, 55] . However, these attributes, such as "big nose", "chubby" and "smiling" in CelebA dataset [26] , can only provide abstract semantic-level guidance on the generation, and the generated faces often lack 3D geometric consistency. Moreover, it is often much harder to obtain low-level geometric annotations beyond semantic labels for direct 3D supervision. Recently, researchers have attempted to incorporate 3D priors from parametric face models, such as 3D Morphable Models (3DMMs) [2, 38] , into StyleGAN-based synthesis models, allowing for more precise 3D control over the generated images, including facial expressions and head poses [8, 52, 39] . Despite their impressive image quality, these models still tend to produce 3D inconsistent faces under large expression and pose variations due to the lack of a 3D representation, as shown in Fig. 1 .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper14
- NDC-Scene: Boost Monocular 3D Semantic Scene Completion in Normalized Device Coordinates SpaceJiawei Yao, Chuming Li, Keqiang Sun, Yingjie Cai 等ICCV 2023 · 被引用 150 次
- Generalizable and Animatable Gaussian Head AvatarXuangeng Chu, Tatsuya HaradaNeurIPS 2024 · 被引用 115 次
- ObjectSDF++: Improved Object-Compositional Neural Implicit SurfacesQianyi Wu, Kaisiyuan Wang, Kejie Li, Jianmin Zheng 等ICCV 2023 · 被引用 44 次
- NaviNeRF: NeRF-based 3D Representation Disentanglement by Latent Semantic NavigationBaao Xie, Bohan Li, Zequn Zhang, Junting Dong 等ICCV 2023 · 被引用 12 次
- Generalizable One-shot 3D Neural Head AvatarXueting Li, Shalini De Mello, Sifei Liu, Koki Nagano 等NeurIPS 2023 · 被引用 12 次
它引用的顶会 Paper29
- Implicit Neural Representations with Periodic Activation FunctionsVincent Sitzmann, Julien N. P. Martel, Alexander W. Bergman, David B. Lindell 等NeurIPS 2020 · 被引用 4,008 次
- Alias-Free Generative Adversarial NetworksTero Karras, Miika Aittala, Samuli Laine, Erik Härkönen 等NeurIPS 2021 · 被引用 2,126 次
- GANSpace: Discovering Interpretable GAN ControlsErik Härkönen, Aaron Hertzmann, Jaakko Lehtinen, Sylvain ParisNeurIPS 2020 · 被引用 1,049 次
- GRAF: Generative Radiance Fields for 3D-Aware Image SynthesisKatja Schwarz, Yiyi Liao, Michael Niemeyer, Andreas GeigerNeurIPS 2020 · 被引用 1,001 次
- Efficient Geometry-aware 3D Generative Adversarial NetworksEric R. Chan, Connor Z. Lin, Matthew A. Chan, Koki Nagano 等CVPR 2022 · 被引用 984 次
相关 Paper
- StyleRig: Rigging StyleGAN for 3D Control Over Portrait ImagesAyush Tewari, Mohamed A. Elgharib, Gaurav Bharaj, Florian Bernard 等CVPR 2020
- MOST-GAN: 3D Morphable StyleGAN for Disentangled Face Image ManipulationSafa C. Medin, Bernhard Egger, Anoop Cherian, Ye Wang 等AAAI 2022 · 被引用 38 次
- Text-Conditional Attribute Alignment Across Latent Spaces for 3D Controllable Face Image SynthesisFeifan Xu, Rui Li, Si Wu, Yong Xu 等CVPR 2024
- MoRF: Morphable Radiance Fields for Multiview Neural Head ModelingDaoye Wang, Prashanth Chandran, Gaspard Zoss, Derek Bradley 等SIGGRAPH 2022 · 被引用 54 次
- GAN-Control: Explicitly Controllable GANsAlon Shoshan, Nadav Bhonker, Igor Kviatkovsky, Gérard G. MedioniICCV 2021 · 被引用 151 次
