Disentangled3D: Learning a 3D Generative Model with Disentangled Geometry and Appearance from Monocular Images
Ayush Tewari, Mallikarjun B. R., Xingang Pan, Ohad Fried, Maneesh Agrawala, Christian Theobalt
Abstract
Learning 3D generative models from a dataset of monocular images enables self-supervised 3D reasoning and controllable synthesis. State-of-the-art 3D generative models are GANs which use neural 3D volumetric representations for synthesis. Images are synthesized by rendering the volumes from a given camera. These models can disentangle the 3D scene from the camera viewpoint in any generated image. However, most models do not disentangle other factors of image formation, such as geometry and appearance. In this paper, we design a 3D GAN which can learn a disentangled model of objects, just from monocular observations. Our model can disentangle the geometry and appearance variations in the scene, i.e., we can independently sample from the geometry and appearance spaces of the generative model. This is achieved using a novel non-rigid deformable scene formulation. A 3D volume which represents an object instance is computed as a non-rigidly deformed canonical 3D volume. Our method learns the canonical volume, as well as its deformations, jointly during training. This formulation also helps us improve the disentanglement between the 3D scene and the camera viewpoints using a novel pose regularization loss defined on the 3D deformation field. In addition, we further model the inverse deformations, enabling the computation of dense correspondences between images generated by our model. Finally, we design an approach to embed real images onto the latent space of our disentangled generative model, enabling editing of real images.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3f8925b9-63eb-4825-8e54-aa8b4386e816Cited by top-tier papers16
- Drag Your GAN: Interactive Point-based Manipulation on the Generative Image ManifoldXingang Pan, Ayush Tewari, Thomas Leimkühler, Lingjie Liu et al.SIGGRAPH 2023 · 206 citations
- VQ3D: Learning a 3D-Aware Generative Model on ImageNetKyle Sargent, Jing Yu Koh, Han Zhang, Huiwen Chang et al.ICCV 2023 · 33 citations
- Improving 3D-aware Image Synthesis with A Geometry-aware DiscriminatorZifan Shi, Yinghao Xu, Yujun Shen, Deli Zhao et al.NeurIPS 2022 · 23 citations
- DeformToon3d: Deformable Neural Radiance Fields for 3D ToonificationJunzhe Zhang, Yushi Lan, Shuai Yang, Fangzhou Hong et al.ICCV 2023 · 16 citations
- NAISR: A 3D Neural Additive Model for Interpretable Shape RepresentationYining Jiao, Carlton J. Zdanski, Julia S. Kimbell, Andrew Prince et al.ICLR 2024 · 7 citations
Builds on24
- Alias-Free Generative Adversarial NetworksTero Karras, Miika Aittala, Samuli Laine, Erik Härkönen et al.NeurIPS 2021 · 2,126 citations
- GRAF: Generative Radiance Fields for 3D-Aware Image SynthesisKatja Schwarz, Yiyi Liao, Michael Niemeyer, Andreas GeigerNeurIPS 2020 · 1,001 citations
- StyleNeRF: A Style-based 3D Aware Generator for High-resolution Image SynthesisJiatao Gu, Lingjie Liu, Peng Wang, Christian TheobaltICLR 2022 · 622 citations
- Non-Rigid Neural Radiance Fields: Reconstruction and Novel View Synthesis of a Dynamic Scene From Monocular VideoEdgar Tretschk, Ayush Tewari, Vladislav Golyanik, Michael Zollhöfer et al.ICCV 2021 · 617 citations
- Editing Conditional Radiance FieldsSteven Liu, Xiuming Zhang, Zhoutong Zhang, Richard Zhang et al.ICCV 2021 · 297 citations
Related papers
- Towards Unsupervised Learning of Generative Models for 3D Controllable Image SynthesisYiyi Liao, Katja Schwarz, Lars M. Mescheder, Andreas GeigerCVPR 2020
- HoloGAN: Unsupervised Learning of 3D Representations From Natural ImagesThu Nguyen-Phuoc, Chuan Li, Lucas Theis, Christian Richardt et al.ICCV 2019 · 98 citations
- Generative Neural Articulated Radiance FieldsAlexander W. Bergman, Petr Kellnhofer, Wang Yifan, Eric R. Chan et al.NeurIPS 2022 · 144 citations
- BlockGAN: Learning 3D Object-aware Scene Representations from Unlabelled ImagesThu Nguyen-Phuoc, Christian Richardt, Long Mai, Yong-Liang Yang et al.NeurIPS 2020 · 256 citations
- GIRAFFE: Representing Scenes As Compositional Generative Neural Feature FieldsMichael Niemeyer, Andreas GeigerCVPR 2021
