Interpretable and Explainable Logical Policies via Neurally Guided Symbolic Abstraction
Quentin Delfosse, Hikaru Shindo, Devendra Singh Dhami, Kristian Kersting
Abstract
The limited priors required by neural networks make them the dominating choice to encode and learn policies using reinforcement learning (RL). However, they are also black-boxes, making it hard to understand the agent's behaviour, especially when working on the image level. Therefore, neuro-symbolic RL aims at creating policies that are interpretable in the first place. Unfortunately, interpretability is not explainability. To achieve both, we introduce Neurally gUided Differentiable loGic policiEs (NUDGE). NUDGE exploits trained neural networkbased agents to guide the search of candidate-weighted logic rules, then uses differentiable logic to train the logic agents. Our experimental evaluation demonstrates that NUDGE agents can induce interpretable and explainable policies while outperforming purely neural ones and showing good flexibility to environments of different initial states and problem sizes. * Equal contribution. † DSD contributed while being with hessian.AI and TU Darmstadt before joining TU. 37th Conference on Neural Information Processing Systems (NeurIPS 2023).
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b3db9340-81ca-43f6-968b-23a0ca066effCited by top-tier papers12
- Interpretable Concept Bottlenecks to Align Reinforcement Learning AgentsQuentin Delfosse, Sebastian Sztwiertnia, Mark Rothermel, Wolfgang Stammer et al.NeurIPS 2024 · 32 citations
- Adaptive Rational Activations to Boost Deep Reinforcement LearningQuentin Delfosse, Patrick Schramowski, Martin Mundt, Alejandro Molina et al.ICLR 2024 · 25 citations
- End-to-End Neuro-Symbolic Reinforcement Learning with Textual ExplanationsLirui Luo, Guoxi Zhang, Hongming Xu, Yaodong Yang et al.ICML 2024 · 18 citations
- Reinforcement Learning Fine-Tuning Enhances Activation Intensity and Diversity in the Internal Circuitry of LLMsHonglin Zhang, Qianyue Hao, Fengli Xu, Yong LiICLR 2026 · 9 citations
- Multimodal LLM-assisted Evolutionary Search for Programmatic Control PoliciesQinglong Hu, Tong Xialiang, Mingxuan Yuan, Fei Liu et al.ICLR 2026 · 7 citations
Builds on8
- Explainable Reinforcement Learning through a Causal LensPrashan Madumal, Tim Miller, Liz Sonenberg, Frank VetereAAAI 2020 · 408 citations
- SPACE: Unsupervised Object-Oriented Scene Representation via Spatial Attention and DecompositionZhixuan Lin, Yi-Fu Wu, Skand Vishwanath Peri, Weihao Sun et al.ICLR 2020 · 276 citations
- Discovering symbolic policies with deep reinforcement learningMikel Landajuela, Brenden K. Petersen, Sookyung Kim, Cláudio P. Santiago et al.ICML 2021 · 118 citations
- Program Guided AgentShao-Hua Sun, Te-Lin Wu, Joseph J. LimICLR 2020 · 63 citations
- Differentiable Inductive Logic Programming for Structured ExamplesHikaru Shindo, Masaaki Nishino, Akihiro YamamotoAAAI 2021 · 40 citations
Related papers
- Neuro-Symbolic Inductive Logic Programming with Logical Neural NetworksPrithviraj Sen, Breno W. S. R. de Carvalho, Ryan Riegel, Alexander G. GrayAAAI 2022 · 82 citations
- Efficient Symbolic Policy Learning with Differentiable Symbolic ExpressionJiaming Guo, Rui Zhang, Shaohui Peng, Qi Yi et al.NeurIPS 2023 · 15 citations
- DeepProofLog: Efficient Proving in Deep Stochastic Logic ProgramsYing Jiao, Rodrigo Castellano Ontiveros, Luc De Raedt, Marco Gori et al.AAAI 2026
- BlendRL: A Framework for Merging Symbolic and Neural Policy LearningHikaru Shindo, Quentin Delfosse, Devendra Singh Dhami, Kristian KerstingICLR 2025
- Learn to Explain Efficiently via Neural Logic Inductive LearningYuan Yang, Le SongICLR 2020 · 83 citations
