Online Bayesian Goal Inference for Boundedly Rational Planning Agents
Tan Zhi-Xuan, Jordyn L. Mann, Tom Silver, Josh Tenenbaum, Vikash Mansinghka
摘要
People routinely infer the goals of others by observing their actions over time. Remarkably, we can do so even when those actions lead to failure, enabling us to assist others when we detect that they might not achieve their goals. How might we endow machines with similar capabilities? Here we present an architecture capable of inferring an agent's goals online from both optimal and non-optimal sequences of actions. Our architecture models agents as boundedly-rational planners that interleave search with execution by replanning, thereby accounting for sub-optimal behavior. These models are specified as probabilistic programs, allowing us to represent and perform efficient Bayesian inference over an agent's goals and internal planning processes. To perform such inference, we develop Sequential Inverse Plan Search (SIPS), a sequential Monte Carlo algorithm that exploits the online replanning assumption of these models, limiting computation by incrementally extending inferred plans as new actions are observed. We present experiments showing that this modeling and inference architecture outperforms Bayesian inverse reinforcement learning baselines, accurately inferring goals from both optimal and non-optimal trajectories involving failure and back-tracking, while generalizing across domains with compositional structure and sparse rewards. Preprint. Under review.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper17
- Predicate Invention for Bilevel PlanningTom Silver, Rohan Chitnis, Nishanth Kumar, Willie McClinton 等AAAI 2023 · 被引用 73 次
- AutoToM: Scaling Model-based Mental Inference via Automated Agent ModelingZhining Zhang, Chuanyang Jin, Mung Yao Jia, Shunchi Zhang 等NeurIPS 2025 · 被引用 30 次
- Inverse Decision Modeling: Learning Interpretable Representations of BehaviorDaniel Jarrett, Alihan Hüyük, Mihaela van der SchaarICML 2021 · 被引用 30 次
- Inverse Optimal Control Adapted to the Noise Characteristics of the Human Sensorimotor SystemMatthias Schultheis, Dominik Straub, Constantin A. RothkopfNeurIPS 2021 · 被引用 25 次
- Robust Neuro-Symbolic Goal and Plan RecognitionLeonardo Amado, Ramon Fraga Pereira, Felipe MeneguzziAAAI 2023 · 被引用 14 次
相关 Paper
- What do you know? Bayesian knowledge inference for navigating agentsMatthias Schultheis, Jana-Sophie Schönfeld, Constantin A. Rothkopf, Heinz KoepplNeurIPS 2025
- Bayes-Adaptive Monte-Carlo Planning and Learning for Goal-Oriented DialoguesYoungsoo Jang, Jongmin Lee, Kee-Eung KimAAAI 2020 · 被引用 22 次
- Procedure Planning in Instructional Videos via Contextual Modeling and Model-based Policy LearningJing Bi, Jiebo Luo, Chenliang XuICCV 2021 · 被引用 64 次
- Inverse Rational Control with Partially Observable Continuous Nonlinear DynamicsMinhae Kwon, Saurabh Daptardar, Paul R. Schrater, Xaq PitkowNeurIPS 2020 · 被引用 46 次
- A Hierarchical Bayesian Approach to Inverse Reinforcement Learning with Symbolic Reward MachinesWeichao Zhou, Wenchao LiICML 2022 · 被引用 15 次
