Goal Recognition as Reinforcement Learning
Leonardo Amado, Reuth Mirsky, Felipe Meneguzzi
Abstract
Most approaches for goal recognition rely on specifications of the possible dynamics of the actor in the environment when pursuing a goal. These specifications suffer from two key issues. First, encoding these dynamics requires careful design by a domain expert, which is often not robust to noise at recognition time. Second, existing approaches often need costly real-time computations to reason about the likelihood of each potential goal. In this paper, we develop a framework that combines model-free reinforcement learning and goal recognition to alleviate the need for careful, manual domain design, and the need for costly online executions. This framework consists of two main stages: Offline learning of policies or utility functions for each potential goal, and online inference. We provide a first instance of this framework using tabular Q-learning for the learning stage, as well as three measures that can be used to perform the inference stage. The resulting instantiation achieves state-of-the-art performance against goal recognizers on standard evaluation domains and superior performance in noisy environments.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext bd0483e9-a152-484e-bcf3-3a645ed6ce7fCited by top-tier papers1
Ask how each one uses itBuilds on1
Related papers
- Information Shaping for Enhanced Goal Recognition of Partially-Informed AgentsSarah Keren, Haifeng Xu, Kofi Kwapong, David C. Parkes et al.AAAI 2020 · 14 citations
- An LP-Based Approach for Goal Recognition as PlanningLuísa R. de A. Santos, Felipe Meneguzzi, Ramon Fraga Pereira, André Grahl PereiraAAAI 2021 · 21 citations
- Value-driven Hindsight ModellingArthur Guez, Fabio Viola, Theophane Weber, Lars Buesing et al.NeurIPS 2020 · 12 citations
- Outcome-Driven Reinforcement Learning via Variational InferenceTim G. J. Rudner, Vitchyr Pong, Rowan McAllister, Yarin Gal et al.NeurIPS 2021 · 24 citations
- Learning Value Functions from Undirected State-only ExperienceMatthew Chang, Arjun Gupta, Saurabh GuptaICLR 2022 · 9 citations
