Unleashing the Potential of Multi-modal Foundation Models and Video Diffusion for 4D Dynamic Physical Scene Simulation
Zhuoman Liu, Weicai Ye, Yan Luximon, Pengfei Wan, Di Zhang
摘要
Realistic simulation of dynamic scenes requires accurately capturing diverse material properties and modeling complex object interactions grounded in physical principles. However, existing methods are constrained to basic material types with limited predictable parameters, making them insufficient to represent the complexity of real-world materials. We introduce PhysFlow, a novel approach that leverages multi-modal foundation models and video diffusion to achieve enhanced 4D dynamic scene simulation. Our method utilizes multi-modal models to identify material types and initialize material parameters through image queries, while simultaneously inferring 3D Gaussian splats for detailed scene representation. We further refine these material parameters using video diffusion with a differentiable Material Point Method (MPM) and optical flow guidance rather than render loss or Score Distillation Sampling (SDS) loss. This integrated framework enables accurate prediction and realistic simulation of dynamic interactions in real-world scenarios, advancing both accuracy and flexibility in physics-based simulations. Our code and data are available at https://zhuomanliu.github.io/PhysFlow
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper12
- MoAlign: Motion-Centric Representation Alignment for Video Diffusion ModelsAritra Bhowmik, Denis Korzhenkov, Cees G. M. Snoek, Amir Habibian 等ICLR 2026 · 被引用 15 次
- VoMP: Predicting Volumetric Mechanical Property FieldsRishit Dagli, Donglai Xiang, Vismay Modi, Charles Loop 等ICLR 2026 · 被引用 13 次
- VisionLaw: Inferring Interpretable Intrinsic Dynamics from Visual Observations via Bilevel OptimizationJiajing Lin, Shu Jiang, Qingyuan Zeng, Zhenzhong Wang 等ICLR 2026 · 被引用 4 次
- DiffWind: Physics-Informed Differentiable Modeling of Wind-Driven Object DynamicsYuanhang Lei, Boming Zhao, Zesong Yang, Xingxuan Li 等ICLR 2026 · 被引用 3 次
- FieryGS: In-the-Wild Fire Synthesis with Physics-Integrated Gaussian SplattingQianfan Shen, Ningxiao Tao, Qiyu Dai, Tianle Chen 等ICLR 2026 · 被引用 3 次
它引用的顶会 Paper15
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 被引用 5,687 次
- Instant neural graphics primitives with a multiresolution hash encodingThomas Müller, Alex Evans, Christoph Schied, Alexander KellerSIGGRAPH 2022 · 被引用 4,089 次
- Mip-NeRF 360: Unbounded Anti-Aliased Neural Radiance FieldsJonathan T. Barron, Ben Mildenhall, Dor Verbin, Pratul P. Srinivasan 等CVPR 2022 · 被引用 1,603 次
- Nerfstudio: A Modular Framework for Neural Radiance Field DevelopmentMatthew Tancik, Ethan Weber, Evonne Ng, Ruilong Li 等SIGGRAPH 2023 · 被引用 592 次
- MotionCtrl: A Unified and Flexible Motion Controller for Video GenerationZhouxia Wang, Ziyang Yuan, Xintao Wang, Yaowei Li 等SIGGRAPH 2024 · 被引用 123 次
相关 Paper
- Phys4DRT: Physics-based 4D Generation for Real-Time Interaction with Time-Frequency SupervisionYuntian Xiao, Shoulong Zhang, Zihang Zhang, Jiahao Cui 等ACM MM 2025
- DreamPhysics: Learning Physics-Based 3D Dynamics with Video Diffusion PriorsTianyu Huang, Haoze Zhang, Yihan Zeng, Zhilu Zhang 等AAAI 2025 · 被引用 21 次
- PhysGM: Large Physical Gaussian Model for Feed-Forward 4D SynthesisChunji Lv, Zequn Chen, Donglin Di, Weinan Zhang 等CVPR 2026 · 被引用 9 次
- Gaussian-Flow: 4D Reconstruction with Dynamic 3D Gaussian ParticleYoutian Lin, Zuozhuo Dai, Siyu Zhu, Yao YaoCVPR 2024
- Physics-Informed Deformable Gaussian Splatting: Towards Unified Constitutive Laws for Time-Evolving Material FieldHaoqin Hong, Ding Fan, Fubin Dou, Zhi-Li Zhou 等AAAI 2026 · 被引用 1 次
