Physics-as-Inverse-Graphics: Unsupervised Physical Parameter Estimation from Video
Miguel Jaques, Michael Burke, Timothy M. Hospedales
摘要
We propose a model that is able to perform physical parameter estimation of systems from video, where the differential equations governing the scene dynamics are known, but labeled states or objects are not available. Existing physical scene understanding methods require either object state supervision, or do not integrate with differentiable physics to learn interpretable system parameters and states. We address this problem through a physics-as-inverse-graphics approach that brings together vision-as-inverse-graphics and differentiable physics engines, where objects and explicit state and velocity representations are discovered by the model. This framework allows us to perform long term extrapolative video prediction, as well as vision-based model-predictive control. Our approach significantly outperforms related unsupervised methods in long-term future frame prediction of systems with interacting objects (such as ball-spring or 3-body gravitational systems), due to its ability to build dynamics into the model as an inductive bias. We further show the value of this tight vision-physics integration by demonstrating data-efficient learning of vision-actuated model-based control for a pendulum system. We also show that the controller's interpretability provides unique capabilities in goal-driven control and physical reasoning for zero-data adaptation.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper23
- Physics-Integrated Variational Autoencoders for Robust and Interpretable Generative ModelingNaoya Takeishi, Alexandros KalousisNeurIPS 2021 · 被引用 88 次
- Unsupervised Learning of Lagrangian Dynamics from Images for Prediction and ControlYaofeng Desmond Zhong, Naomi Ehrich LeonardNeurIPS 2020 · 被引用 49 次
- NewtonGen: Physics-consistent and Controllable Text-to-Video Generation via Neural Newtonian DynamicsYu Yuan, Xijun Wang, Tharindu Wickremasinghe, Zeeshan Nadir 等ICLR 2026 · 被引用 46 次
- Learning Physics Constrained Dynamics Using AutoencodersTsung-Yen Yang, Justinian Rosca, Karthik Narasimhan, Peter J. RamadgeNeurIPS 2022 · 被引用 39 次
- -SfT: Shape-from-Template with a Physics-Based Deformation ModelNavami Kairanda, Edith Tretschk, Mohamed A. Elgharib, Christian Theobalt 等CVPR 2022 · 被引用 20 次
相关 Paper
- Dynamic Visual Reasoning by Learning Differentiable Physics Models from Video and LanguageMingyu Ding, Zhenfang Chen, Tao Du, Ping Luo 等NeurIPS 2021 · 被引用 90 次
- Seeing the Wind from a Falling LeafZhiyuan Gao, Jiageng Mao, Hong-Xing Yu, Haozhe Lou 等NeurIPS 2025 · 被引用 10 次
- gradSim: Differentiable simulation for system identification and visuomotor controlJ. Krishna Murthy, Miles Macklin, Florian Golemo, Vikram Voleti 等ICLR 2021 · 被引用 130 次
- Visual Grounding of Learned Physical ModelsYunzhu Li, Toru Lin, Kexin Yi, Daniel Bear 等ICML 2020 · 被引用 88 次
- Latent Intuitive Physics: Learning to Transfer Hidden Physics from A 3D VideoXiangming Zhu, Huayu Deng, Haochen Yuan, Yunbo Wang 等ICLR 2024 · 被引用 5 次
