Synthesis of Search Heuristics for Temporal Planning via Reinforcement Learning
Andrea Micheli, Alessandro Valentini
Abstract
Automated temporal planning is the problem of synthesizing, starting from a model of a system, a course of actions to achieve a desired goal when temporal constraints, such as deadlines, are present in the problem. Despite considerable successes in the literature, scalability is still a severe limitation for existing planners, especially when confronted with real-world, industrial scenarios.
In this paper, we aim at exploiting recent advances in reinforcement learning, for the synthesis of heuristics for temporal planning. Starting from a set of problems of interest for a specific domain, we use a customized reinforcement learning algorithm to construct a value function that is able to estimate the expected reward for as many problems as possible. We use a reward schema that captures the semantics of the temporal planning problem and we show how the value function can be transformed in a planning heuristic for a semi-symbolic heuristic search exploration of the planning model. We show on two case-studies how this method can widen the reach of current temporal planners with encouraging results.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 72ac8c99-bb24-46d3-aa43-fd7b8e384e6cCited by top-tier papers3
- Graph Learning for Numeric PlanningDillon Z. Chen, Sylvie ThiébauxNeurIPS 2024 · 8 citations
- Temporal Task and Motion Planning with Metric Time for Multiple Object NavigationElisa Tosello, Alessandro Valentini, Andrea MicheliAAAI 2025 · 2 citations
- Automatic Selection of Macro-Events for Heuristic-Search Temporal PlanningAlessandro La Farciola, Alessandro Valentini, Andrea MicheliAAAI 2025
Builds on2
- Temporal Planning with Intermediate Conditions and EffectsAlessandro Valentini, Andrea Micheli, Alessandro CimattiAAAI 2020 · 27 citations
- Decidability and Complexity of Action-Based Temporal Planning over Dense TimeNicola Gigante, Andrea Micheli, Angelo Montanari, Enrico ScalaAAAI 2020 · 20 citations
Related papers
- Synthesis from Satisficing and Temporal GoalsSuguman Bansal, Lydia E. Kavraki, Moshe Y. Vardi, Andrew M. WellsAAAI 2022 · 6 citations
- Expressive Optimal Temporal Planning via Optimization Modulo TheoryStefan Panjkovic, Andrea MicheliAAAI 2023 · 8 citations
- Temporal-Logic-Based Reward Shaping for Continuing Reinforcement Learning TasksYuqian Jiang, Suda Bharadwaj, Bo Wu, Rishi Shah et al.AAAI 2021 · 54 citations
- Event-Triggered and Time-Triggered Duration Calculus for Model-Free Reinforcement LearningKalyani Dole, Ashutosh Gupta, John Komp, Shankaranarayanan Krishna et al.RTSS 2021 · 3 citations
- Deductive Synthesis of Reinforcement Learning Agents for Infinite Horizon TasksYuning Wang, He ZhuCAV 2025
