Mixed Neural Voxels for Fast Multi-view Video Synthesis
Feng Wang, Sinan Tan, Xinghang Li, Zeyue Tian, Yafei Song, Huaping Liu
摘要
Synthesizing high-fidelity videos from real-world multi-view input is challenging due to the complexities of real-world environments and high-dynamic movements. Previous works based on neural radiance fields have demonstrated high-quality reconstructions of dynamic scenes. However, training such models on real-world scenes is time-consuming, usually taking days or weeks. In this paper, we present a novel method named MixVoxels to efficiently represent dynamic scenes, enabling fast training and rendering speed. The proposed MixVoxels represents the 4D dynamic scenes as a mixture of static and dynamic voxels and processes them with different networks. In this way, the computation of the required modalities for static voxels can be processed by a lightweight model, which essentially reduces the amount of computation as many daily dynamic scenes are dominated by static backgrounds. To distinguish the two kinds of voxels, we propose a novel variation field to estimate the temporal variance of each voxel. For the dynamic representations, we design an inner product time query method to efficiently query multiple time steps, which is essential to recover the high-dynamic movements. As a result, with 15 minutes of training for dynamic scenes with inputs of 300-frame videos, MixVoxels achieves better PSNR than previous methods. For rendering, MixVoxels can render a novel view video with 1K resolution at 37 fps. Codes and trained models are available at https://github.com/fengres/mixvoxels.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper66
- Real-time Photorealistic Dynamic Scene Representation and Rendering with 4D Gaussian SplattingZeyu Yang, Hongye Yang, Zijie Pan, Li ZhangICLR 2024 · 被引用 529 次
- 4D Gaussian Splatting for Real-Time Dynamic Scene RenderingGuanjun Wu, Taoran Yi, Jiemin Fang, Lingxi Xie 等CVPR 2024 · 被引用 513 次
- Efficient Region-Aware Neural Radiance Fields for High-Fidelity Talking Portrait SynthesisJiahe Li, Jiawei Zhang, Xiao Bai, Jun Zhou 等ICCV 2023 · 被引用 127 次
- Fully Explicit Dynamic Gaussian SplattingJunoh Lee, Changyeon Won, Hyunjun Jung, Inhwan Bae 等NeurIPS 2024 · 被引用 92 次
- MotionGS: Exploring Explicit Motion Guidance for Deformable 3D Gaussian SplattingRuijie Zhu, Yanzhe Liang, Hanzhi Chang, Jiacheng Deng 等NeurIPS 2024 · 被引用 87 次
它引用的顶会 Paper26
- Instant neural graphics primitives with a multiresolution hash encodingThomas Müller, Alex Evans, Christoph Schied, Alexander KellerSIGGRAPH 2022 · 被引用 4,089 次
- Fourier Features Let Networks Learn High Frequency Functions in Low Dimensional DomainsMatthew Tancik, Pratul P. Srinivasan, Ben Mildenhall, Sara Fridovich-Keil 等NeurIPS 2020 · 被引用 4,036 次
- Mip-NeRF: A Multiscale Representation for Anti-Aliasing Neural Radiance FieldsJonathan T. Barron, Ben Mildenhall, Matthew Tancik, Peter Hedman 等ICCV 2021 · 被引用 2,700 次
- Mip-NeRF 360: Unbounded Anti-Aliased Neural Radiance FieldsJonathan T. Barron, Ben Mildenhall, Dor Verbin, Pratul P. Srinivasan 等CVPR 2022 · 被引用 1,603 次
- Neural Sparse Voxel FieldsLingjie Liu, Jiatao Gu, Kyaw Zaw Lin, Tat-Seng Chua 等NeurIPS 2020 · 被引用 1,535 次
相关 Paper
- Ced-NeRF: A Compact and Efficient Method for Dynamic Neural Radiance FieldsYoutian LinAAAI 2024 · 被引用 2 次
- Learning Neural Volumetric Representations of Dynamic Humans in MinutesChen Geng, Sida Peng, Zhen Xu, Hujun Bao 等CVPR 2023
- DeVRF: Fast Deformable Voxel Radiance Fields for Dynamic ScenesJiawei Liu, Yan-Pei Cao, Weijia Mao, Wenqiao Zhang 等NeurIPS 2022 · 被引用 151 次
- Neural 3D Video Synthesis from Multi-view VideoTianye Li, Mira Slavcheva, Michael Zollhöfer, Simon Green 等CVPR 2022 · 被引用 324 次
- Streaming Radiance Fields for 3D Video SynthesisLingzhi Li, Zhen Shen, Zhongshu Wang, Li Shen 等NeurIPS 2022 · 被引用 155 次
