EVA3D: Compositional 3D Human Generation from 2D Image Collections
Fangzhou Hong, Zhaoxi Chen, Yushi Lan, Liang Pan, Ziwei Liu
摘要
Inverse graphics aims to recover 3D models from 2D observations. Utilizing differentiable rendering, recent 3D-aware generative models have shown impressive results of rigid object generation using 2D images. However, it remains challenging to generate articulated objects, like human bodies, due to their complexity and diversity in poses and appearances. In this work, we propose, EVA3D, an unconditional 3D human generative model learned from 2D image collections only. EVA3D can sample 3D humans with detailed geometry and render high-quality images (up to 512x256) without bells and whistles (e.g. super resolution). At the core of EVA3D is a compositional human NeRF representation, which divides the human body into local parts. Each part is represented by an individual volume. This compositional representation enables 1) inherent human priors, 2) adaptive allocation of network parameters, 3) efficient training and rendering. Moreover, to accommodate for the characteristics of sparse 2D human image collections (e.g. imbalanced pose distribution), we propose a pose-guided sampling strategy for better GAN learning. Extensive experiments validate that EVA3D achieves state-of-the-art 3D human generation performance regarding both geometry and texture quality. Notably, EVA3D demonstrates great potential and scalability to "inverse-graphics" diverse human bodies with a clean framework.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper55
- SparseNeRF: Distilling Depth Ranking for Few-shot Novel View SynthesisGuangcong Wang, Zhaoxi Chen, Chen Change Loy, Ziwei LiuICCV 2023 · 被引用 309 次
- DreamPose: Fashion Image-to-Video Synthesis via Stable DiffusionJohanna Suvi Karras, Aleksander Holynski, Ting-Chun Wang, Ira Kemelmacher-ShlizermanICCV 2023 · 被引用 224 次
- HIFA: High-fidelity Text-to-3D Generation with Advanced Diffusion GuidanceJunzhe Zhu, Peiye Zhuang, Sanmi KoyejoICLR 2024 · 被引用 113 次
- MagicAnimate: Temporally Consistent Human Image Animation using Diffusion ModelZhongcong Xu, Jianfeng Zhang, Jun Hao Liew, Hanshu Yan 等CVPR 2024 · 被引用 106 次
- AG3D: Learning to Generate 3D Avatars from 2D Image CollectionsZijian Dong, Xu Chen, Jinlong Yang, Michael J. Black 等ICCV 2023 · 被引用 76 次
它引用的顶会 Paper36
- Implicit Neural Representations with Periodic Activation FunctionsVincent Sitzmann, Julien N. P. Martel, Alexander W. Bergman, David B. Lindell 等NeurIPS 2020 · 被引用 4,008 次
- AMASS: Archive of Motion Capture As Surface ShapesNaureen Mahmood, Nima Ghorbani, Nikolaus F. Troje, Gerard Pons-Moll 等ICCV 2019 · 被引用 1,784 次
- Learning to Reconstruct 3D Human Pose and Shape via Model-Fitting in the LoopNikos Kolotouros, Georgios Pavlakos, Michael J. Black, Kostas DaniilidisICCV 2019 · 被引用 1,139 次
- Implicit Geometric Regularization for Learning ShapesAmos Gropp, Lior Yariv, Niv Haim, Matan Atzmon 等ICML 2020 · 被引用 1,001 次
- GRAF: Generative Radiance Fields for 3D-Aware Image SynthesisKatja Schwarz, Yiyi Liao, Michael Niemeyer, Andreas GeigerNeurIPS 2020 · 被引用 1,001 次
相关 Paper
- 3DHumanGAN: 3D-Aware Human Image Generation with 3D Pose MappingZhuoqian Yang, Shikai Li, Wayne Wu, Bo DaiICCV 2023 · 被引用 19 次
- Learning Compositional Radiance Fields of Dynamic Human HeadsZiyan Wang, Timur M. Bagautdinov, Stephen Lombardi, Tomas Simon 等CVPR 2021
- DiffHuman: Probabilistic Photorealistic 3D Reconstruction of HumansAkash Sengupta, Thiemo Alldieck, Nikos Kolotouros, Enric Corona 等CVPR 2024 · 被引用 10 次
- GeneMAN: Generalizable Single-Image 3D Human Reconstruction from Multi-Source Human DataWentao Wang, Hang Ye, Fangzhou Hong, Xue Yang 等NeurIPS 2025 · 被引用 6 次
- Image GANs meet Differentiable Rendering for Inverse Graphics and Interpretable 3D Neural RenderingYuxuan Zhang, Wenzheng Chen, Huan Ling, Jun Gao 等ICLR 2021 · 被引用 140 次
