NVFi: Neural Velocity Fields for 3D Physics Learning from Dynamic Videos
Jinxi Li, Ziyang Song, Bo Yang
Abstract
In this paper, we aim to model 3D scene dynamics from multi-view videos. Unlike the majority of existing works which usually focus on the common task of novel view synthesis within the training time period, we propose to simultaneously learn the geometry, appearance, and physical velocity of 3D scenes only from video frames, such that multiple desirable applications can be supported, including future frame extrapolation, unsupervised 3D semantic scene decomposition, and dynamic motion transfer. Our method consists of three major components, 1) the keyframe dynamic radiance field, 2) the interframe velocity field, and 3) a joint keyframe and interframe optimization module which is the core of our framework to effectively train both networks. To validate our method, we further introduce two dynamic 3D datasets: 1) Dynamic Object dataset, and 2) Dynamic Indoor Scene dataset. We conduct extensive experiments on multiple datasets, demonstrating the superior performance of our method over all baselines, particularly in the critical tasks of future frame extrapolation and unsupervised 3D semantic scene decomposition. Our code and data are available at https://github.com/vLAR-group/NVFi
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 5343fe9d-83b6-46e1-bfd4-6671e38c9e43Cited by top-tier papers8
- PURE: Prompt Evolution with Graph ODE for Out-of-distribution Fluid Dynamics ModelingHao Wu, Changhu Wang, Fan Xu, Jinbao Xue et al.NeurIPS 2024 · 23 citations
- ODE-GS: Latent ODEs for Dynamic Scene Extrapolation with 3D Gaussian SplattingDaniel Wang, Patrick Rim, Tian Tian, Dong Lao et al.ICLR 2026 · 12 citations
- Seeing the Wind from a Falling LeafZhiyuan Gao, Jiageng Mao, Hong-Xing Yu, Haozhe Lou et al.NeurIPS 2025 · 10 citations
- ParticleGS: Learning Neural Gaussian Particle Dynamics from Videos for Prior-free Physical Motion ExtrapolationJinsheng Quan, Qiaowei Miao, Yichao Xu, Zizhuo Lin et al.CVPR 2026 · 5 citations
- Multi-Object System Identification from VideosChunjiang Liu, Xiaoyuan Wang, Qingran Lin, Albert Xiao et al.ICLR 2026 · 2 citations
Builds on29
- Fourier Features Let Networks Learn High Frequency Functions in Low Dimensional DomainsMatthew Tancik, Pratul P. Srinivasan, Ben Mildenhall, Sara Fridovich-Keil et al.NeurIPS 2020 · 4,036 citations
- Implicit Neural Representations with Periodic Activation FunctionsVincent Sitzmann, Julien N. P. Martel, Alexander W. Bergman, David B. Lindell et al.NeurIPS 2020 · 4,008 citations
- Dynamic View Synthesis from Dynamic Monocular VideoChen Gao, Ayush Saraf, Johannes Kopf, Jia-Bin HuangICCV 2021 · 522 citations
- Neural Unsigned Distance Fields for Implicit Function LearningJulian Chibane, Aymen Mir, Gerard Pons-MollNeurIPS 2020 · 415 citations
- HumanNeRF: Free-viewpoint Rendering of Moving People from Monocular VideoChung-Yi Weng, Brian Curless, Pratul P. Srinivasan, Jonathan T. Barron et al.CVPR 2022 · 411 citations
Related papers
- TRACE: Learning 3D Gaussian Physical Dynamics from Multi-View VideosJinxi Li, Ziyang Song, Bo YangICCV 2025 · 3 citations
- FreeGave: 3D Physics Learning from Dynamic Videos by Gaussian VelocityJinxi Li, Ziyang Song, Siyuan Zhou, Bo YangCVPR 2025
- Unsupervised Volumetric AnimationAliaksandr Siarohin, Willi Menapace, Ivan Skorokhodov, Kyle Olszewski et al.CVPR 2023
- Neural 3D Video Synthesis from Multi-view VideoTianye Li, Mira Slavcheva, Michael Zollhöfer, Simon Green et al.CVPR 2022 · 324 citations
- STaR: Self-Supervised Tracking and Reconstruction of Rigid Objects in Motion With Neural RenderingWentao Yuan, Zhaoyang Lv, Tanner Schmidt, Steven LovegroveCVPR 2021
