MonoNeRF: Learning a Generalizable Dynamic Radiance Field from Monocular Videos
Fengrui Tian, Shaoyi Du, Yueqi Duan
摘要
In this paper, we target at the problem of learning a generalizable dynamic radiance field from monocular videos. Different from most existing NeRF methods that are based on multiple views, monocular videos only contain one view at each timestamp, thereby suffering from ambiguity along the view direction in estimating point features and scene flows. Previous studies such as DynNeRF disambiguate point features by positional encoding, which is not transferable and severely limits the generalization ability. As a result, these methods have to train one independent model for each scene and suffer from heavy computational costs when applying to increasing monocular videos in real-world applications. To address this, We propose MonoNeRF to simultaneously learn point features and scene flows with point trajectory and feature correspondence constraints across frames. More specifically, we learn an implicit velocity field to estimate point trajectory from temporal features with Neural ODE, which is followed by a flow-based feature aggregation module to obtain spatial features along the point trajectory. We jointly optimize temporal and spatial features in an end-to-end manner. Experiments show that our MonoNeRF is able to learn from multiple scenes and support new applications such as scene editing, unseen frame synthesis, and fast novel scene adaptation. Codes are available at https://github.com/tianfr/MonoNeRF.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper35
- 4D Gaussian Splatting for Real-Time Dynamic Scene RenderingGuanjun Wu, Taoran Yi, Jiemin Fang, Lingxi Xie 等CVPR 2024 · 被引用 513 次
- L4GM: Large 4D Gaussian Reconstruction ModelJiawei Ren, Cheng Xie, Ashkan Mirzaei, Hanxue Liang 等NeurIPS 2024 · 被引用 173 次
- 4D-Rotor Gaussian Splatting: Towards Efficient Novel View Synthesis for Dynamic ScenesYuanxing Duan, Fangyin Wei, Qiyu Dai, Yuhang He 等SIGGRAPH 2024 · 被引用 121 次
- Feed-Forward Bullet-Time Reconstruction of Dynamic Scenes from Monocular VideosHanxue Liang, Jiawei Ren, Ashkan Mirzaei, Antonio Torralba 等NeurIPS 2025 · 被引用 52 次
- GFlow: Recovering 4D World from Monocular VideoShizun Wang, Xingyi Yang, Qiuhong Shen, Zhenxiang Jiang 等AAAI 2025 · 被引用 47 次
它引用的顶会 Paper28
- SlowFast Networks for Video RecognitionChristoph Feichtenhofer, Haoqi Fan, Jitendra Malik, Kaiming HeICCV 2019 · 被引用 4,104 次
- NeuS: Learning Neural Implicit Surfaces by Volume Rendering for Multi-view ReconstructionPeng Wang, Lingjie Liu, Yuan Liu, Christian Theobalt 等NeurIPS 2021 · 被引用 2,500 次
- Neural Sparse Voxel FieldsLingjie Liu, Jiatao Gu, Kyaw Zaw Lin, Tat-Seng Chua 等NeurIPS 2020 · 被引用 1,535 次
- MVSNeRF: Fast Generalizable Radiance Field Reconstruction from Multi-View StereoAnpei Chen, Zexiang Xu, Fuqiang Zhao, Xiaoshuai Zhang 等ICCV 2021 · 被引用 1,024 次
- GRAF: Generative Radiance Fields for 3D-Aware Image SynthesisKatja Schwarz, Yiyi Liao, Michael Niemeyer, Andreas GeigerNeurIPS 2020 · 被引用 1,001 次
相关 Paper
- MonoNeRF: Learning Generalizable NeRFs from Monocular Videos without Camera PosesYang Fu, Ishan Misra, Xiaolong WangICML 2023 · 被引用 13 次
- DynPoint: Dynamic Neural Point For View SynthesisKaichen Zhou, Jia-Xing Zhong, Sangyun Shin, Kai Lu 等NeurIPS 2023 · 被引用 46 次
- Semantic Flow: Learning Semantic Fields of Dynamic Scenes from Monocular VideosFengrui Tian, Yueqi Duan, Angtian Wang, Jianfei Guo 等ICLR 2024 · 被引用 7 次
- Dynamic View Synthesis from Dynamic Monocular VideoChen Gao, Ayush Saraf, Johannes Kopf, Jia-Bin HuangICCV 2021 · 被引用 522 次
- DyBluRF: Dynamic Neural Radiance Fields from Blurry Monocular VideoHuiqiang Sun, Xingyi Li, Liao Shen, Xinyi Ye 等CVPR 2024 · 被引用 5 次
