Autonomous Capability Assessment of Sequential Decision-Making Systems in Stochastic Settings
Pulkit Verma, Rushang Karia, Siddharth Srivastava
摘要
It is essential for users to understand what their AI systems can and can't do in order to use them safely. However, the problem of enabling users to assess AI systems with sequential decision-making (SDM) capabilities is relatively understudied. This paper presents a new approach for modeling the capabilities of black-box AI systems that can plan and act, along with the possible effects and requirements for executing those capabilities in stochastic settings. We present an active-learning approach that can effectively interact with a black-box SDM system and learn an interpretable probabilistic model describing its capabilities. Theoretical analysis of the approach identifies the conditions under which the learning process is guaranteed to converge to the correct model of the agent; empirical evaluations on different agents and simulated scenarios show that this approach is few-shot generalizable and can effectively describe the capabilities of arbitrary black-box SDM agents in a sample-efficient manner.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper10
- Online Bayesian Goal Inference for Boundedly Rational Planning AgentsTan Zhi-Xuan, Jordyn L. Mann, Tom Silver, Josh Tenenbaum 等NeurIPS 2020 · 被引用 122 次
- Creativity of AI: Automatic Symbolic Option Discovery for Facilitating Deep Reinforcement LearningMu Jin, Zhihao Ma, Kebing Jin, Hankz Hankui Zhuo 等AAAI 2022 · 被引用 49 次
- PDSketch: Integrated Domain Programming, Learning, and PlanningJiayuan Mao, Tomás Lozano-Pérez, Josh Tenenbaum, Leslie Pack KaelblingNeurIPS 2022 · 被引用 40 次
- Bridging the Gap: Providing Post-Hoc Symbolic Explanations for Sequential Decision-Making Problems with Inscrutable RepresentationsSarath Sreedharan, Utkarsh Soni, Mudit Verma, Siddharth Srivastava 等ICLR 2022 · 被引用 39 次
- GLIB: Efficient Exploration for Relational Model-Based Reinforcement Learning via Goal-Literal BabblingRohan Chitnis, Tom Silver, Joshua B. Tenenbaum, Leslie Pack Kaelbling 等AAAI 2021 · 被引用 38 次
相关 Paper
- Differential Assessment of Black-Box AI AgentsRashmeet Kaur Nayyar, Pulkit Verma, Siddharth SrivastavaAAAI 2022 · 被引用 19 次
- Asking the Right Questions: Learning Interpretable Action Models Through Query AnsweringPulkit Verma, Shashank Rao Marpally, Siddharth SrivastavaAAAI 2021 · 被引用 33 次
- Learning Interpretable Temporal Properties from Positive Examples OnlyRajarshi Roy, Jean-Raphaël Gaglione, Nasim Baharisangari, Daniel Neider 等AAAI 2023 · 被引用 21 次
- Active Statistical InferenceTijana Zrnic, Emmanuel J. CandèsICML 2024 · 被引用 34 次
- Teaching Humans When to Defer to a Classifier via ExemplarsHussein Mozannar, Arvind Satyanarayan, David A. SontagAAAI 2022 · 被引用 49 次
