I-PHYRE: Interactive Physical Reasoning
Shiqian Li, Kewen Wu, Chi Zhang, Yixin Zhu
摘要
Current evaluation protocols predominantly assess physical reasoning in stationary scenes, creating a gap in evaluating agents' abilities to interact with dynamic events. While contemporary methods allow agents to modify initial scene configurations and observe consequences, they lack the capability to interact with events in real time. To address this, we introduce I-PHYRE, a framework that challenges agents to simultaneously exhibit intuitive physical reasoning, multi-step planning, and in-situ intervention. Here, intuitive physical reasoning refers to a quick, approximate understanding of physics to address complex problems; multi-step denotes the need for extensive sequence planning in I-PHYRE, considering each intervention can significantly alter subsequent choices; and in-situ implies the necessity for timely object manipulation within a scene, where minor timing deviations can result in task failure. We formulate four game splits to scrutinize agents' learning and generalization of essential principles of interactive physical reasoning, fostering learning through interaction with representative scenarios. Our exploration involves three planning strategies and examines several supervised and reinforcement agents' zero-shot generalization proficiency on I-PHYRE. The outcomes highlight a notable gap between existing learning algorithms and human performance, emphasizing the imperative for more research in enhancing agents with interactive physical reasoning capabilities. The environment and baselines will be made publicly available.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- SparseDFF: Sparse-View Feature Distillation for One-Shot Dexterous ManipulationQianxu Wang, Haotong Zhang, Congyue Deng, Yang You 等ICLR 2024 · 被引用 36 次
- Learning Interactive World Model for Object-Centric Reinforcement LearningFan Feng, Phillip Lippe, Sara MagliacaneNeurIPS 2025 · 被引用 13 次
- DeepPhy: Benchmarking Agentic VLMs on Physical ReasoningXinrun Xu, Pi Bu, Ye Wang, Börje F. Karlsson 等AAAI 2026 · 被引用 6 次
- Learning Physics-Grounded 4D Dynamics with Neural Gaussian Force FieldsShiqian Li, Ruihong Shen, Junfeng Ni, Chang Pan 等ICLR 2026 · 被引用 5 次
- Neural Force Field: Few-shot Learning of Generalized Physical ReasoningShiqian Li, Ruihong Shen, Yaoyu Tao, Chi Zhang 等ICLR 2026 · 被引用 1 次
它引用的顶会 Paper7
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Decision Transformer: Reinforcement Learning via Sequence ModelingLili Chen, Kevin Lu, Aravind Rajeswaran, Kimin Lee 等NeurIPS 2021 · 被引用 2,557 次
- CLEVRER: Collision Events for Video Representation and ReasoningKexin Yi, Chuang Gan, Yunzhu Li, Pushmeet Kohli 等ICLR 2020 · 被引用 584 次
- ComPhy: Compositional Physical Reasoning of Objects and Events from VideosZhenfang Chen, Kexin Yi, Yunzhu Li, Mingyu Ding 等ICLR 2022 · 被引用 67 次
- Learning Long-term Visual Dynamics with Region Proposal Interaction NetworksHaozhi Qi, Xiaolong Wang, Deepak Pathak, Yi Ma 等ICLR 2021 · 被引用 63 次
相关 Paper
- Causal-PIK: Causality-based Physical Reasoning with a Physics-Informed KernelCarlota Parés-Morlans, Michelle Yi, Claire Chen, Sarah A. Wu 等ICML 2025
- IPR-1: Interactive Physical ReasonerMingyu Zhang, Lifeng Zhuo, Tianxi Tan, Guocan Xie 等CVPR 2026 · 被引用 2 次
- On the Learning Mechanisms in Physical ReasoningShiqian Li, Kewen Wu, Chi Zhang, Yixin ZhuNeurIPS 2022 · 被引用 21 次
- Real-Time Reasoning Agents in Evolving EnvironmentsYule Wen, Yixin Ye, Yanzhe Zhang, Diyi Yang 等ICLR 2026 · 被引用 10 次
- Beyond Static Vision: Scene Dynamic Field Unlocks Intuitive Physics Understanding in Multi-modal Large Language ModelsNanxi Li, Xiang Wang, Yuanjie Chen, Haode Zhang 等ICLR 2026
