SLAP: Shortcut Learning for Abstract Planning
Y. Isabel Liu, Bowen Li, Benjamin Eysenbach, Tom Silver
Abstract
Long-horizon decision-making with sparse rewards and continuous states and actions remains a fundamental challenge in AI and robotics. Task and motion planning (TAMP) is a model-based framework that addresses this challenge by planning hierarchically with abstract actions (options). These options are manually defined, limiting the agent to behaviors that we as human engineers know how to program (pick, place, move). In this work, we propose Shortcut Learning for Abstract Planning (SLAP), a method that leverages existing TAMP options to automatically discover new ones. Our key idea is to use model-free reinforcement learning (RL) to learn shortcuts in the abstract planning graph induced by the existing options in TAMP. Without any additional assumptions or inputs, shortcut learning leads to shorter solutions than pure planning, and higher task success rates than flat and hierarchical RL. Qualitatively, SLAP discovers dynamic physical improvisations (e.g., slap, wiggle, wipe) that differ significantly from the manually-defined ones. In experiments in four simulated robotic environments, we show that SLAP solves and generalizes to a wide range of tasks, reducing overall plan lengths by over 50% and consistently outperforming planning and RL baselines.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Related papers
- Discovering State and Action Abstractions for Generalized Task and Motion PlanningAidan Curtis, Tom Silver, Joshua B. Tenenbaum, Tomás Lozano-Pérez et al.AAAI 2022 · 36 citations
- Flexible and Efficient Long-Range Planning Through Curious ExplorationAidan Curtis, Minjian Xin, Dilip Arumugam, Kevin T. Feigelis et al.ICML 2020 · 7 citations
- Consciousness-Inspired Spatio-Temporal Abstractions for Better Generalization in Reinforcement LearningHarry Zhao, Safa Alver, Harm van Seijen, Romain Laroche et al.ICLR 2024 · 5 citations
- One Demo Is All It Takes: Planning Domain Derivation with LLMs from A Single DemonstrationJinbang Huang, Yixin Xiao, Zhanguang Zhang, Mark Coates et al.ICLR 2026 · 9 citations
- Learning Geometric Reasoning Networks For Robot Task And Motion PlanningSmail Ait Bouhsain, Rachid Alami, Thierry SiméonICLR 2025
