Autonomous Capability Assessment of Sequential Decision-Making Systems in Stochastic Settings
Pulkit Verma, Rushang Karia, Siddharth Srivastava
Abstract
It is essential for users to understand what their AI systems can and can't do in order to use them safely. However, the problem of enabling users to assess AI systems with sequential decision-making (SDM) capabilities is relatively understudied. This paper presents a new approach for modeling the capabilities of black-box AI systems that can plan and act, along with the possible effects and requirements for executing those capabilities in stochastic settings. We present an active-learning approach that can effectively interact with a black-box SDM system and learn an interpretable probabilistic model describing its capabilities. Theoretical analysis of the approach identifies the conditions under which the learning process is guaranteed to converge to the correct model of the agent; empirical evaluations on different agents and simulated scenarios show that this approach is few-shot generalizable and can effectively describe the capabilities of arbitrary black-box SDM agents in a sample-efficient manner.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c164bda8-47d1-4f7a-afcd-ddd8d8f81166Builds on10
- Online Bayesian Goal Inference for Boundedly Rational Planning AgentsTan Zhi-Xuan, Jordyn L. Mann, Tom Silver, Josh Tenenbaum et al.NeurIPS 2020 · 122 citations
- Creativity of AI: Automatic Symbolic Option Discovery for Facilitating Deep Reinforcement LearningMu Jin, Zhihao Ma, Kebing Jin, Hankz Hankui Zhuo et al.AAAI 2022 · 49 citations
- PDSketch: Integrated Domain Programming, Learning, and PlanningJiayuan Mao, Tomás Lozano-Pérez, Josh Tenenbaum, Leslie Pack KaelblingNeurIPS 2022 · 40 citations
- Bridging the Gap: Providing Post-Hoc Symbolic Explanations for Sequential Decision-Making Problems with Inscrutable RepresentationsSarath Sreedharan, Utkarsh Soni, Mudit Verma, Siddharth Srivastava et al.ICLR 2022 · 39 citations
- GLIB: Efficient Exploration for Relational Model-Based Reinforcement Learning via Goal-Literal BabblingRohan Chitnis, Tom Silver, Joshua B. Tenenbaum, Leslie Pack Kaelbling et al.AAAI 2021 · 38 citations
Related papers
- Differential Assessment of Black-Box AI AgentsRashmeet Kaur Nayyar, Pulkit Verma, Siddharth SrivastavaAAAI 2022 · 19 citations
- Asking the Right Questions: Learning Interpretable Action Models Through Query AnsweringPulkit Verma, Shashank Rao Marpally, Siddharth SrivastavaAAAI 2021 · 33 citations
- Learning Interpretable Temporal Properties from Positive Examples OnlyRajarshi Roy, Jean-Raphaël Gaglione, Nasim Baharisangari, Daniel Neider et al.AAAI 2023 · 21 citations
- Active Statistical InferenceTijana Zrnic, Emmanuel J. CandèsICML 2024 · 34 citations
- Teaching Humans When to Defer to a Classifier via ExemplarsHussein Mozannar, Arvind Satyanarayan, David A. SontagAAAI 2022 · 49 citations
