Symbolic Network: Generalized Neural Policies for Relational MDPs
Sankalp Garg, Aniket Bajpai, Mausam
Abstract
A Relational Markov Decision Process (RMDP) is a first-order representation to express all instances of a single probabilistic planning domain with possibly unbounded number of objects. Early work in RMDPs outputs generalized (instance-independent) first-order policies or value functions as a means to solve all instances of a domain at once. Unfortunately, this line of work met with limited success due to inherent limitations of the representation space used in such policies or value functions. Can neural models provide the missing link by easily representing more complex generalized policies, thus making them effective on all instances of a given domain? We present SymNet, the first neural approach for solving RMDPs that are expressed in the probabilistic planning language of RDDL. SymNet trains a set of shared parameters for an RDDL domain using training instances from that domain. For each instance, SymNet first converts it to an instance graph and then uses relational neural models to compute node embeddings. It then scores each ground action as a function over the first-order action symbols and node embeddings related to the action. Given a new test instance from the same domain, SymNet architecture with pre-trained parameters scores each ground action and chooses the best action. This can be accomplished in a single forward pass without any retraining on the test instance, thus implicitly representing a neural generalized policy for the whole domain. Our experiments on nine RDDL domains from IPPC demonstrate that SymNet policies are significantly better than random and sometimes even more effective than training a state-of-the-art deep reactive policy from scratch.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8d166e2c-4809-4d4e-a3ba-c67ec13a32e7Cited by top-tier papers8
- Learning General Planning Policies from Small Examples Without SupervisionGuillem Francès, Blai Bonet, Hector GeffnerAAAI 2021 · 44 citations
- Learning Domain-Independent Heuristics for Grounded and Lifted PlanningDillon Ze Chen, Sylvie Thiébaux, Felipe W. TrevizanAAAI 2024 · 29 citations
- A Solver-free Framework for Scalable Learning in Neural ILP ArchitecturesYatin Nandwani, Rishabh Ranjan, Mausam, Parag SinglaNeurIPS 2022 · 13 citations
- Sample-Efficient Iterative Lower Bound Optimization of Deep Reactive Policies for Planning in Continuous MDPsSiow Meng Low, Akshat Kumar, Scott SannerAAAI 2022 · 3 citations
- Neural Models for Output-Space Invariance in Combinatorial ProblemsYatin Nandwani, Vidit Jain, Mausam, Parag SinglaICLR 2022 · 3 citations
Related papers
- Graph Neural Network Based Action Ranking for PlanningRajesh Mangannavar, Stefan Lee, Alan Fern, Prasad TadepalliNeurIPS 2025 · 3 citations
- What Planning Problems Can A Relational Neural Network Solve?Jiayuan Mao, Tomás Lozano-Pérez, Joshua B. Tenenbaum, Leslie Pack KaelblingNeurIPS 2023 · 13 citations
- Learning More Expressive General Policies for Classical Planning DomainsSimon Ståhlberg, Blai Bonet, Hector GeffnerAAAI 2025
- Generating Programmatic Referring Expressions via Program SynthesisJiani Huang, Calvin Smith, Osbert Bastani, Rishabh Singh et al.ICML 2020 · 11 citations
- Generalized Planning for the Abstraction and Reasoning CorpusChao Lei, Nir Lipovetzky, Krista A. EhingerAAAI 2024 · 13 citations
