Probabilistic inverse optimal control for non-linear partially observable systems disentangles perceptual uncertainty and behavioral costs
Dominik Straub, Matthias Schultheis, Heinz Koeppl, Constantin A. Rothkopf
摘要
Inverse optimal control can be used to characterize behavior in sequential decisionmaking tasks. Most existing work, however, is limited to fully observable or linear systems, or requires the action signals to be known. Here, we introduce a probabilistic approach to inverse optimal control for partially observable stochastic non-linear systems with unobserved action signals, which unifies previous approaches to inverse optimal control with maximum causal entropy formulations. Using an explicit model of the noise characteristics of the sensory and motor systems of the agent in conjunction with local linearization techniques, we derive an approximate likelihood function for the model parameters, which can be computed within a single forward pass. We present quantitative evaluations on stochastic and partially observable versions of two classic control tasks and two human behavioral tasks. Importantly, we show that our method can disentangle perceptual factors and behavioral costs despite the fact that epistemic and pragmatic actions are intertwined in sequential decision-making under uncertainty, such as in active sensing and active learning. The proposed method has broad applicability, ranging from imitation learning to sensorimotor neuroscience. * equal contribution Preprint. Under review.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper7
- Dream to Control: Learning Behaviors by Latent ImaginationDanijar Hafner, Timothy P. Lillicrap, Jimmy Ba, Mohammad NorouziICLR 2020 · 被引用 1,852 次
- Efficient and Modular Implicit DifferentiationMathieu Blondel, Quentin Berthet, Marco Cuturi, Roy Frostig 等NeurIPS 2022 · 被引用 386 次
- IQ-Learn: Inverse soft-Q Learning for ImitationDivyansh Garg, Shuvam Chakraborty, Chris Cundy, Jiaming Song 等NeurIPS 2021 · 被引用 271 次
- Inverse Rational Control with Partially Observable Continuous Nonlinear DynamicsMinhae Kwon, Saurabh Daptardar, Paul R. Schrater, Xaq PitkowNeurIPS 2020 · 被引用 46 次
- Inferring learning rules from animal decision-makingZoe Ashwood, Nicholas A. Roy, Ji Hyun Bak, Jonathan W. PillowNeurIPS 2020 · 被引用 33 次
相关 Paper
- Inverse Optimal Control Adapted to the Noise Characteristics of the Human Sensorimotor SystemMatthias Schultheis, Dominik Straub, Constantin A. RothkopfNeurIPS 2021 · 被引用 25 次
- Inverse decision-making using neural amortized Bayesian actorsDominik Straub, Tobias F. Niehues, Jan Peters, Constantin A. RothkopfICLR 2025
- Maximum Likelihood Constraint Inference for Inverse Reinforcement LearningDexter R. R. Scobee, S. Shankar SastryICLR 2020 · 被引用 74 次
- Stochastic Optimal Control and Estimation with Multiplicative and Internal NoiseFrancesco Damiani, Akiyuki Anzai, Jan Drugowitsch, Gregory C. DeAngelis 等NeurIPS 2024 · 被引用 2 次
- Inverse Reinforcement Learning with Explicit Policy EstimatesNavyata Sanghvi, Shinnosuke Usami, Mohit Sharma, Joachim Groeger 等AAAI 2021 · 被引用 7 次
