Make Encoder Great Again in 3D GAN Inversion through Geometry and Occlusion-Aware Encoding
Ziyang Yuan, Yiming Zhu, Yu Li, Hongyu Liu, Chun Yuan
摘要
3D GAN inversion aims to achieve high reconstruction fidelity and reasonable 3D geometry simultaneously from a single image input. However, existing 3D GAN inversion methods rely on time-consuming optimization for each individual case. In this work, we introduce a novel encoder-based inversion framework based on EG3D, one of the most widely-used 3D GAN models. We leverage the inherent properties of EG3D’s latent space to design a discriminator and a background depth regularization. This enables us to train a geometry-aware encoder capable of converting the input image into corresponding latent code. Additionally, we explore the feature space of EG3D and develop an adaptive refinement stage that improves the representation ability of features in EG3D to enhance the recovery of fine-grained textural details. Finally, we propose an occlusion-aware fusion operation to prevent distortion in unobserved regions. Our method achieves impressive results comparable to optimization-based methods while operating up to 500 times faster. Our framework is well-suited for applications such as semantic editing.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper24
- VOODOO 3D: Volumetric Portrait Disentanglement for One-Shot 3D Head ReenactmentPhong Tran, Egor Zakharov, Long-Nhat Ho, Anh Tuan Tran 等CVPR 2024 · 被引用 15 次
- Lite2Relight: 3D-aware Single Image Portrait RelightingPramod Rao, Gereon Fox, Abhimitra Meka, Mallikarjun B. R. 等SIGGRAPH 2024 · 被引用 14 次
- Generalizable One-shot 3D Neural Head AvatarXueting Li, Shalini De Mello, Sifei Liu, Koki Nagano 等NeurIPS 2023 · 被引用 12 次
- Dual Encoder GAN Inversion for High-Fidelity 3D Head Reconstruction from Single ImagesBahri Batuhan Bilecen, Ahmet Berke Gökmen, Aysegul DundarNeurIPS 2024 · 被引用 11 次
- Holo-Relighting: Controllable Volumetric Portrait Relighting from a Single ImageYiqun Mei, Yu Zeng, He Zhang, Zhixin Shu 等CVPR 2024 · 被引用 11 次
它引用的顶会 Paper28
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- On the Variance of the Adaptive Learning Rate and BeyondLiyuan Liu, Haoming Jiang, Pengcheng He, Weizhu Chen 等ICLR 2020 · 被引用 2,210 次
- StyleCLIP: Text-Driven Manipulation of StyleGAN ImageryOr Patashnik, Zongze Wu, Eli Shechtman, Daniel Cohen-Or 等ICCV 2021 · 被引用 1,437 次
- Image2StyleGAN: How to Embed Images Into the StyleGAN Latent Space?Rameen Abdal, Yipeng Qin, Peter WonkaICCV 2019 · 被引用 1,195 次
- Efficient Geometry-aware 3D Generative Adversarial NetworksEric R. Chan, Connor Z. Lin, Matthew A. Chan, Koki Nagano 等CVPR 2022 · 被引用 984 次
相关 Paper
- High-fidelity 3D GAN Inversion by Pseudo-multi-view OptimizationJiaxin Xie, Hao Ouyang, Jingtan Piao, Chenyang Lei 等CVPR 2023
- Self-Supervised Geometry-Aware Encoder for Style-Based 3D GAN InversionYushi Lan, Xuyi Meng, Shuai Yang, Chen Change Loy 等CVPR 2023
- Reference-Based 3D-Aware Image Editing with TriplanesBahri Batuhan Bilecen, Yigit Yalin, Ning Yu, Aysegul DundarCVPR 2025
- 3D GAN Inversion with Facial Symmetry PriorFei Yin, Yong Zhang, Xuan Wang, Tengfei Wang 等CVPR 2023
- InvertAvatar: Incremental GAN Inversion for Generalized Head AvatarsXiaochen Zhao, Jingxiang Sun, Lizhen Wang, Jinli Suo 等SIGGRAPH 2024 · 被引用 10 次
