Learning Probably Approximately Complete and Safe Action Models for Stochastic Worlds
Brendan Juba, Roni Stern
Abstract
We consider the problem of learning action models for planning in unknown stochastic environments that can be defined using the Probabilistic Planning Domain Description Language (PPDDL). As input, we are given a set of previously executed trajectories, and the main challenge is to learn an action model that has a similar goal achievement probability to the policies used to create these trajectories. To this end, we introduce a variant of PPDDL in which there is uncertainty about the transition probabilities, specified by an interval for each factor that contains the respective true transition probabilities. Then, we present SAM+, an algorithm that learns such an imprecise-PPDDL environment model. SAM+ has a polynomial time and sample complexity, and guarantees that with high probability, the true environment is indeed captured by the defined intervals. We prove that the action model SAM+ outputs has a goal achievement probability that is almost as good or better than that of the policies used to produced the training trajectories. Then, we show how to produce a PPDDL model based on this imprecise-PPDDL model that has similar properties.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext db40af80-e179-4b49-9565-5655b3e5a63eCited by top-tier papers4
- Autonomous Capability Assessment of Sequential Decision-Making Systems in Stochastic SettingsPulkit Verma, Rushang Karia, Siddharth SrivastavaNeurIPS 2023 · 14 citations
- Optimistic Exploration in Reinforcement Learning Using Symbolic Model EstimatesSarath Sreedharan, Michael KatzNeurIPS 2023 · 12 citations
- Learning Safe Action Models with Partial ObservabilityHai S. Le, Brendan Juba, Roni SternAAAI 2024 · 7 citations
- Learning Safe Numeric Action ModelsArgaman Mordoch, Brendan Juba, Roni SternAAAI 2023 · 6 citations
Builds on2
Related papers
- On-line Learning of Planning Domains from Sensor Data in PAL: Scaling up to Large State SpacesLeonardo Lamanna, Alfonso Emilio Gerevini, Alessandro Saetti, Luciano Serafini et al.AAAI 2021 · 10 citations
- Planning with Uncertain Action ModelsFrancesco Percassi, Alessandro Saetti, Enrico ScalaAAAI 2026
- On the Limit of Language Models as Planning FormalizersCassie Huang, Li ZhangACL 2025
- Sampling-Based Robust Control of Autonomous Systems with Non-Gaussian NoiseThom S. Badings, Alessandro Abate, Nils Jansen, David Parker et al.AAAI 2022 · 33 citations
- Planning with Abstract Learned Models While Learning Transferable SubtasksJohn Winder, Stephanie Milani, Matthew Landen, Erebus Oh et al.AAAI 2020 · 10 citations
