Visual Grounding of Learned Physical Models
Yunzhu Li, Toru Lin, Kexin Yi, Daniel Bear, Daniel Yamins, Jiajun Wu, Joshua B. Tenenbaum, Antonio Torralba
摘要
Humans intuitively recognize objects' physical properties and predict their motion, even when the objects are engaged in complicated interactions. The abilities to perform physical reasoning and to adapt to new environments, while intrinsic to humans, remain challenging to state-of-the-art computational models. In this work, we present a neural model that simultaneously reasons about physics and makes future predictions based on visual and dynamics priors. The visual prior predicts a particle-based representation of the system from visual observations. An inference module operates on those particles, predicting and refining estimates of particle locations, object states, and physical parameters, subject to the constraints imposed by the dynamics prior, which we refer to as visual grounding. We demonstrate the effectiveness of our method in environments involving rigid objects, deformable materials, and fluids. Experiments show that our model can infer the physical properties within a few observations, which allows the model to quickly adapt to unseen scenarios and make accurate predictions into the future.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper21
- Causal Discovery in Physical Systems from VideosYunzhu Li, Antonio Torralba, Anima Anandkumar, Dieter Fox 等NeurIPS 2020 · 被引用 133 次
- gradSim: Differentiable simulation for system identification and visuomotor controlJ. Krishna Murthy, Miles Macklin, Florian Golemo, Vikram Voleti 等ICLR 2021 · 被引用 130 次
- Dynamic Visual Reasoning by Learning Differentiable Physics Models from Video and LanguageMingyu Ding, Zhenfang Chen, Tao Du, Ping Luo 等NeurIPS 2021 · 被引用 90 次
- Learning Physical Graph Representations from Visual ScenesDaniel Bear, Chaofei Fan, Damian Mrowca, Yunzhu Li 等NeurIPS 2020 · 被引用 88 次
- Learning Physical Dynamics with Subequivariant Graph Neural NetworksJiaqi Han, Wenbing Huang, Hengbo Ma, Jiachen Li 等NeurIPS 2022 · 被引用 72 次
它引用的顶会 Paper4
- Learning to Simulate Complex Physics with Graph NetworksAlvaro Sanchez-Gonzalez, Jonathan Godwin, Tobias Pfaff, Rex Ying 等ICML 2020 · 被引用 1,439 次
- CLEVRER: Collision Events for Video Representation and ReasoningKexin Yi, Chuang Gan, Yunzhu Li, Pushmeet Kohli 等ICLR 2020 · 被引用 584 次
- Lagrangian Fluid Simulation with Continuous ConvolutionsBenjamin Ummenhofer, Lukas Prantl, Nils Thuerey, Vladlen KoltunICLR 2020 · 被引用 211 次
- Learning Compositional Koopman Operators for Model-Based ControlYunzhu Li, Hao He, Jiajun Wu, Dina Katabi 等ICLR 2020 · 被引用 135 次
相关 Paper
- 3D-IntPhys: Towards More Generalized 3D-grounded Visual Intuitive Physics under Challenging ScenesHaotian Xue, Antonio Torralba, Josh Tenenbaum, Dan Yamins 等NeurIPS 2023 · 被引用 19 次
- SlotPi: Physics-informed Object-centric Reasoning ModelsJian Li, Han Wan, Ning Lin, Yu-Liang Zhan 等KDD 2025
- NeuroFluid: Fluid Dynamics Grounding with Particle-Driven Neural Radiance FieldsShanyan Guan, Huayu Deng, Yunbo Wang, Xiaokang YangICML 2022 · 被引用 51 次
- Physics-as-Inverse-Graphics: Unsupervised Physical Parameter Estimation from VideoMiguel Jaques, Michael Burke, Timothy M. HospedalesICLR 2020 · 被引用 58 次
- Hierarchical Relational InferenceAleksandar Stanic, Sjoerd van Steenkiste, Jürgen SchmidhuberAAAI 2021 · 被引用 17 次
