PV3D: A 3D Generative Model for Portrait Video Generation
Eric Zhongcong Xu, Jianfeng Zhang, Jun Hao Liew, Wenqing Zhang, Song Bai, Jiashi Feng, Mike Zheng Shou
摘要
Recent advances in generative adversarial networks (GANs) have demonstrated the capabilities of generating stunning photo-realistic portrait images. While some prior works have applied such image GANs to unconditional 2D portrait video generation and static 3D portrait synthesis, there are few works successfully extending GANs for generating 3D-aware portrait videos. In this work, we propose PV3D, the first generative framework that can synthesize multi-view consistent portrait videos. Specifically, our method extends the recent static 3D-aware image GAN to the video domain by generalizing the 3D implicit neural representation to model the spatio-temporal space. To introduce motion dynamics to the generation process, we develop a motion generator by stacking multiple motion layers to generate motion features via modulated convolution. To alleviate motion ambiguities caused by camera/human motions, we propose a simple yet effective camera condition strategy for PV3D, enabling both temporal and multi-view consistent video generation. Moreover, PV3D introduces two discriminators for regularizing the spatial and temporal domains to ensure the plausibility of the generated portrait videos. These elaborated designs enable PV3D to generate 3D-aware motion-plausible portrait videos with high-quality appearance and geometry, significantly outperforming prior works. As a result, PV3D is able to support many downstream applications such as animating static portraits and view-consistent video motion editing. Code and models are released at https://showlab.github.io/pv3d.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- LAM: Large Avatar Model for One-shot Animatable Gaussian HeadYisheng He, Xiaodong Gu, Xiaodan Ye, Chao Xu 等SIGGRAPH 2025 · 被引用 14 次
- OrthoPlanes: A Novel Representation for Better 3D-Awareness of GANsHonglin He, Zhuoqian Yang, Shikai Li, Bo Dai 等ICCV 2023 · 被引用 9 次
- Real-Time 3D-Aware Portrait Video RelightingZiqi Cai, Kaiwen Jiang, Shu-Yu Chen, Yu-Kun Lai 等CVPR 2024
- CAP4D: Creating Animatable 4D Portrait Avatars with Morphable Multi-View Diffusion ModelsFelix Taubner, Ruihang Zhang, Mathieu Tuli, David B. LindellCVPR 2025
- Feed-Forward One-Shot Animatable Textured Mesh Avatar ReconstructionYisheng HeCVPR 2026
它引用的顶会 Paper27
- Implicit Neural Representations with Periodic Activation FunctionsVincent Sitzmann, Julien N. P. Martel, Alexander W. Bergman, David B. Lindell 等NeurIPS 2020 · 被引用 4,008 次
- Alias-Free Generative Adversarial NetworksTero Karras, Miika Aittala, Samuli Laine, Erik Härkönen 等NeurIPS 2021 · 被引用 2,126 次
- Image2StyleGAN: How to Embed Images Into the StyleGAN Latent Space?Rameen Abdal, Yipeng Qin, Peter WonkaICCV 2019 · 被引用 1,195 次
- GRAF: Generative Radiance Fields for 3D-Aware Image SynthesisKatja Schwarz, Yiyi Liao, Michael Niemeyer, Andreas GeigerNeurIPS 2020 · 被引用 1,001 次
- Efficient Geometry-aware 3D Generative Adversarial NetworksEric R. Chan, Connor Z. Lin, Matthew A. Chan, Koki Nagano 等CVPR 2022 · 被引用 984 次
相关 Paper
- AniFaceGAN: Animatable 3D-Aware Face Image Generation for Video AvatarsYue Wu, Yu Deng, Jiaolong Yang, Fangyun Wei 等NeurIPS 2022 · 被引用 77 次
- Pi-GAN: Periodic Implicit Generative Adversarial Networks for 3D-Aware Image SynthesisEric R. Chan, Marco Monteiro, Petr Kellnhofer, Jiajun Wu 等CVPR 2021
- 3DHumanGAN: 3D-Aware Human Image Generation with 3D Pose MappingZhuoqian Yang, Shikai Li, Wayne Wu, Bo DaiICCV 2023 · 被引用 19 次
- Next3D: Generative Neural Texture Rasterization for 3D-Aware Head AvatarsJingxiang Sun, Xuan Wang, Lizhen Wang, Xiaoyu Li 等CVPR 2023
- Multi-View Consistent Generative Adversarial Networks for 3D-aware Image SynthesisXuanmeng Zhang, Zhedong Zheng, Daiheng Gao, Bang Zhang 等CVPR 2022 · 被引用 37 次
