On-line Learning of Planning Domains from Sensor Data in PAL: Scaling up to Large State Spaces
Leonardo Lamanna, Alfonso Emilio Gerevini, Alessandro Saetti, Luciano Serafini, Paolo Traverso
摘要
We propose an approach to learn an extensional representation of a discrete deterministic planning domain from observations in a continuous space navigated by the agent actions. This is achieved through the use of a perception function providing the likelihood of a real-value observation being in a given state of the planning domain after executing an action. The agent learns an extensional representation of the domain (the set of states, the transitions from states to states caused by actions) and the perception function on-line, while it acts for accomplishing its task. In order to provide a practical approach that can scale up to large state spaces, a “draft” intensional (PDDL-based) model of the planning domain is used to guide the exploration of the environment and learn the states and state transitions. The proposed approach uses a novel algorithm to (i) construct the extensional representation of the domain by interleaving symbolic planning in the PDDL intensional representation and search in the state transition graph of the extensional representation; (ii) incrementally refine the intensional representation taking into account information about the actions that the agent cannot execute. An experimental analysis shows that the novel approach can scale up to large state spaces, thus overcoming the limits in scalability of the previous work.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Online Learning of Reusable Abstract Models for Object Goal NavigationTommaso Campari, Leonardo Lamanna, Paolo Traverso, Luciano Serafini 等CVPR 2022 · 被引用 19 次
- Differential Assessment of Black-Box AI AgentsRashmeet Kaur Nayyar, Pulkit Verma, Siddharth SrivastavaAAAI 2022 · 被引用 19 次
相关 Paper
- Planning for Learning Object PropertiesLeonardo Lamanna, Luciano Serafini, Mohamadreza Faridghasemnia, Alessandro Saffiotti 等AAAI 2023 · 被引用 12 次
- Learning Probably Approximately Complete and Safe Action Models for Stochastic WorldsBrendan Juba, Roni SternAAAI 2022 · 被引用 18 次
- Learning Safe Action Models with Partial ObservabilityHai S. Le, Brendan Juba, Roni SternAAAI 2024 · 被引用 7 次
- Predicate Invention for Bilevel PlanningTom Silver, Rohan Chitnis, Nishanth Kumar, Willie McClinton 等AAAI 2023 · 被引用 73 次
- PALMER: Perception - Action Loop with Memory for Long-Horizon PlanningOnur Beker, Mohammad Mohammadi, Amir ZamirNeurIPS 2022 · 被引用 6 次
