Lifting 2D StyleGAN for 3D-Aware Face Generation
Yichun Shi, Divyansh Aggarwal, Anil K. Jain
摘要
In this study, we present our results and experience during replicating the paper titled "Lifting 2D StyleGAN for 3D-Aware Face Generation" (1). This work proposes a model, called LiftedGAN, that disentangles the latent space of StyleGAN2 (2) into texture, shape, viewpoint, lighting components and utilizes those components to render novel synthetic images. This approach claims to enable the ability of manipulating viewpoint and lighting components separately without altering other features of the image. We have trained the proposed model in PyTorch (3), and have conducted all experiments presented in the original work. Thereafter, we have written the evaluation code from scratch. Our re-implementation enables us to better compare different models inferring on the same latent vector input. We were able to reproduce most of the results presented in the original paper both qualitatively and quantitatively. Scope of Reproducibility In the scope of this study, we aim to reproduce all of the qualitative and quantitative results of LiftedGAN, including the ablation study, on FFHQ (4) and AFHQ Cat (5) datasets. Additionally, we further extend the experiments presented in the original work by testing the proposed approach on CelebA (6) dataset. Methodology We have adopted the source code for training from the author's repository. We have written the evaluation scripts from scratch in PyTorch to test the original and reproduced weights on the same latent vector. Our experiments have been completed on a single Nvidia Quadro RTX 6000 in 1 day for each, and it requires ∼11GB GPU memory for training. Results We have achieved to reproduce the results qualitatively and quantitatively on a large scale. We also validated the generalization ability of the model by training and testing it on CelebA dataset. Although our experimental results are not identical with the original paper, they are consistent and validates the claims made by the original work. What was easy The paper is well-written. The main components of the LiftedGAN was open-source, and implemented in PyTorch, which facilitated our reproduction study. What was difficult 3D evaluation and reconstruction scripts were not available in the official repository. Also, there were some missing implementation details to reproduce some results in the original work. Communication with original authors We were in contact with the authors since the beginning of the challenge. We could not achieve to reproduce 3D evaluation and reconstruction parts, fortunately, the authors swiftly answered our questions regarding the topic.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper33
- Efficient Geometry-aware 3D Generative Adversarial NetworksEric R. Chan, Connor Z. Lin, Matthew A. Chan, Koki Nagano 等CVPR 2022 · 被引用 984 次
- GRAM: Generative Radiance Manifolds for 3D-Aware Image GenerationYu Deng, Jiaolong Yang, Jianfeng Xiang, Xin TongCVPR 2022 · 被引用 189 次
- VoxGRAF: Fast 3D-Aware Image Synthesis with Sparse Voxel GridsKatja Schwarz, Axel Sauer, Michael Niemeyer, Yiyi Liao 等NeurIPS 2022 · 被引用 185 次
- EpiGRAF: Rethinking training of 3D GANsIvan Skorokhodov, Sergey Tulyakov, Yiqun Wang, Peter WonkaNeurIPS 2022 · 被引用 145 次
- GRAM-HD: 3D-Consistent Image Generation at High Resolution with Generative Radiance ManifoldsJianfeng Xiang, Jiaolong Yang, Yu Deng, Xin TongICCV 2023 · 被引用 95 次
它引用的顶会 Paper4
- HoloGAN: Unsupervised Learning of 3D Representations From Natural ImagesThu Nguyen-Phuoc, Chuan Li, Lucas Theis, Christian Richardt 等ICCV 2019 · 被引用 98 次
- StarGAN v2: Diverse Image Synthesis for Multiple DomainsYunjey Choi, Youngjung Uh, Jaejun Yoo, Jung-Woo HaCVPR 2020
- Disentangled and Controllable Face Image Generation via 3D Imitative-Contrastive LearningYu Deng, Jiaolong Yang, Dong Chen, Fang Wen 等CVPR 2020
- Analyzing and Improving the Image Quality of StyleGANTero Karras, Samuli Laine, Miika Aittala, Janne Hellsten 等CVPR 2020
相关 Paper
- A Latent Transformer for Disentangled Face Editing in Images and VideosXu Yao, Alasdair Newson, Yann Gousseau, Pierre HellierICCV 2021 · 被引用 97 次
- Conceptual and Hierarchical Latent Space Decomposition for Face EditingSavas Özkan, Mete Özay, Tom RobinsonICCV 2023 · 被引用 3 次
- Retrieve in Style: Unsupervised Facial Feature Transfer and RetrievalMin Jin Chong, Wen-Sheng Chu, Abhishek Kumar, David A. ForsythICCV 2021 · 被引用 27 次
- Enhancing Identity-Deformation Disentanglement in StyleGAN for One-Shot Face Video Re-EnactmentQing Chang, Yao-Xiang Ding, Kun ZhouAAAI 2025 · 被引用 3 次
- HyperReenact: One-Shot Reenactment via Jointly Learning to Refine and Retarget FacesStella Bounareli, Christos Tzelepis, Vasileios Argyriou, Ioannis Patras 等ICCV 2023 · 被引用 63 次
