Improving 3D-aware Image Synthesis with A Geometry-aware Discriminator
Zifan Shi, Yinghao Xu, Yujun Shen, Deli Zhao, Qifeng Chen, Dit-Yan Yeung
Abstract
3D-aware image synthesis aims at learning a generative model that can render photo-realistic 2D images while capturing decent underlying 3D shapes. A popular solution is to adopt the generative adversarial network (GAN) and replace the generator with a 3D renderer, where volume rendering with neural radiance field (NeRF) is commonly used. Despite the advancement of synthesis quality, existing methods fail to obtain moderate 3D shapes. We argue that, considering the twoplayer game in the formulation of GANs, only making the generator 3D-aware is not enough. In other words, displacing the generative mechanism only offers the capability, but not the guarantee, of producing 3D-aware images, because the supervision of the generator primarily comes from the discriminator. To address this issue, we propose GeoD through learning a geometry-aware discriminator to improve 3D-aware GANs. Concretely, besides differentiating real and fake samples from the 2D image space, the discriminator is additionally asked to derive the geometry information from the inputs, which is then applied as the guidance of the generator. Such a simple yet effective design facilitates learning substantially more accurate 3D shapes. Extensive experiments on various generator architectures and training datasets verify the superiority of GeoD over state-of-the-art alternatives. Moreover, our approach is registered as a general framework such that a more capable discriminator (i.e., with a third task of novel view synthesis beyond domain classification and geometry extraction) can further assist the generator with a better multi-view consistency. Project page can be found here. * Work was done during an internship under Ant Group. 36th Conference on Neural Information Processing Systems (NeurIPS 2022).
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 14971331-fd67-4e83-948f-9d33c0f8c1f4Cited by top-tier papers8
- SMaRt: Improving GANs with Score Matching RegularityMengfei Xia, Yujun Shen, Ceyuan Yang, Ran Yi et al.ICML 2024 · 9 citations
- CAD : Photorealistic 3D Generation via Adversarial DistillationZiyu Wan, Despoina Paschalidou, Ian Huang, Hongyu Liu et al.CVPR 2024 · 3 citations
- BerfScene: Bev-conditioned Equivariant Radiance Fields for Infinite 3D Scene GenerationQihang Zhang, Yinghao Xu, Yujun Shen, Bo Dai et al.CVPR 2024 · 1 citation
- RODIN: A Generative Model for Sculpting 3D Digital Avatars Using DiffusionTengfei Wang, Bo Zhang, Ting Zhang, Shuyang Gu et al.CVPR 2023
- GLeaD: Improving GANs with A Generator-Leading TaskQingyan Bai, Ceyuan Yang, Yinghao Xu, Xihui Liu et al.CVPR 2023
Builds on21
- Vision Transformers for Dense PredictionRené Ranftl, Alexey Bochkovskiy, Vladlen KoltunICCV 2021 · 2,647 citations
- Alias-Free Generative Adversarial NetworksTero Karras, Miika Aittala, Samuli Laine, Erik Härkönen et al.NeurIPS 2021 · 2,126 citations
- GRAF: Generative Radiance Fields for 3D-Aware Image SynthesisKatja Schwarz, Yiyi Liao, Michael Niemeyer, Andreas GeigerNeurIPS 2020 · 1,001 citations
- Efficient Geometry-aware 3D Generative Adversarial NetworksEric R. Chan, Connor Z. Lin, Matthew A. Chan, Koki Nagano et al.CVPR 2022 · 984 citations
- StyleNeRF: A Style-based 3D Aware Generator for High-resolution Image SynthesisJiatao Gu, Lingjie Liu, Peng Wang, Christian TheobaltICLR 2022 · 622 citations
Related papers
- 3D-aware Image Synthesis via Learning Structural and Textural RepresentationsYinghao Xu, Sida Peng, Ceyuan Yang, Yujun Shen et al.CVPR 2022 · 88 citations
- Learning 3D-Aware Image Synthesis with Unknown Pose DistributionZifan Shi, Yujun Shen, Yinghao Xu, Sida Peng et al.CVPR 2023
- G-NeRF: Geometry-enhanced Novel View Synthesis from Single-View ImagesZixiong Huang, Qi Chen, Libo Sun, Yifan Yang et al.CVPR 2024
- 3D-aware Blending with Generative NeRFsHyunsu Kim, Gayoung Lee, Yunjey Choi, Jin-Hwa Kim et al.ICCV 2023 · 14 citations
- A Shading-Guided Generative Implicit Model for Shape-Accurate 3D-Aware Image SynthesisXingang Pan, Xudong Xu, Chen Change Loy, Christian Theobalt et al.NeurIPS 2021 · 104 citations
