Learning Neural Volumetric Representations of Dynamic Humans in Minutes
Chen Geng, Sida Peng, Zhen Xu, Hujun Bao, Xiaowei Zhou
摘要
This paper addresses the challenge of efficiently reconstructing volumetric videos of dynamic humans from sparse multi-view videos. Some recent works represent a dynamic human as a canonical neural radiance field (NeRF) and a motion field, which are learned from input videos through differentiable rendering. But the per-scene optimization generally requires hours. Other generalizable NeRF models leverage learned prior from datasets to reduce the optimization time by only finetuning on new scenes at the cost of visual fidelity. In this paper, we propose a novel method for learning neural volumetric representations of dynamic humans in minutes with competitive visual quality. Specifically, we define a novel part-based voxelized human representation to better distribute the representational power of the network to different human parts. Furthermore, we propose a novel 2D motion parameterization scheme to increase the convergence rate of deformation field learning. Experiments demonstrate that our model can be learned 100 times faster than previous per-scene optimization methods while being competitive in the rendering quality. Training our model on a 512 × 512 video with 100 frames typically takes about 5 minutes on a single RTX 3090 GPU. The code is available on our project page: https://zju3dv.github.io/instant_nvr.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper32
- 3DGS-Avatar: Animatable Avatars via Deformable 3D Gaussian SplattingZhiyin Qian, Shaofei Wang, Marko Mihajlovic, Andreas Geiger 等CVPR 2024 · 被引用 131 次
- Human Gaussian Splatting: Real-Time Rendering of Animatable AvatarsArthur Moreau, Jifei Song, Helisa Dhamo, Richard Shaw 等CVPR 2024 · 被引用 55 次
- SIFU: Side-view Conditioned Implicit Function for Real-world Usable Clothed Human ReconstructionZechuan Zhang, Zongxin Yang, Yi YangCVPR 2024 · 被引用 44 次
- GoMAvatar: Efficient Animatable Human Modeling from Monocular Video Using Gaussians-on-MeshJing Wen, Xiaoming Zhao, Zhongzheng Ren, Alexander G. Schwing 等CVPR 2024 · 被引用 33 次
- Expressive Gaussian Human Avatars from Monocular RGB VideoHezhen Hu, Zhiwen Fan, Tianhao Wu, Yihan Xi 等NeurIPS 2024 · 被引用 29 次
它引用的顶会 Paper55
- Instant neural graphics primitives with a multiresolution hash encodingThomas Müller, Alex Evans, Christoph Schied, Alexander KellerSIGGRAPH 2022 · 被引用 4,089 次
- Fourier Features Let Networks Learn High Frequency Functions in Low Dimensional DomainsMatthew Tancik, Pratul P. Srinivasan, Ben Mildenhall, Sara Fridovich-Keil 等NeurIPS 2020 · 被引用 4,036 次
- Implicit Neural Representations with Periodic Activation FunctionsVincent Sitzmann, Julien N. P. Martel, Alexander W. Bergman, David B. Lindell 等NeurIPS 2020 · 被引用 4,008 次
- Mip-NeRF: A Multiscale Representation for Anti-Aliasing Neural Radiance FieldsJonathan T. Barron, Ben Mildenhall, Matthew Tancik, Peter Hedman 等ICCV 2021 · 被引用 2,700 次
- NeuS: Learning Neural Implicit Surfaces by Volume Rendering for Multi-view ReconstructionPeng Wang, Lingjie Liu, Yuan Liu, Christian Theobalt 等NeurIPS 2021 · 被引用 2,500 次
相关 Paper
- DeVRF: Fast Deformable Voxel Radiance Fields for Dynamic ScenesJiawei Liu, Yan-Pei Cao, Weijia Mao, Wenqiao Zhang 等NeurIPS 2022 · 被引用 151 次
- Direct Voxel Grid Optimization: Super-fast Convergence for Radiance Fields ReconstructionCheng Sun, Min Sun, Hwann-Tzong ChenCVPR 2022 · 被引用 859 次
- HumanNeRF: Efficiently Generated Human Radiance Field from Sparse InputsFuqiang Zhao, Wei Yang, Jiakai Zhang, Pei Lin 等CVPR 2022 · 被引用 109 次
- Forward Flow for Novel View Synthesis of Dynamic ScenesXiang Guo, Jiadai Sun, Yuchao Dai, Guanying Chen 等ICCV 2023 · 被引用 75 次
- H-NeRF: Neural Radiance Fields for Rendering and Temporal Reconstruction of Humans in MotionHongyi Xu, Thiemo Alldieck, Cristian SminchisescuNeurIPS 2021 · 被引用 225 次
