Right for the Right Reasons: Avoiding Reasoning Shortcuts via Prototypical Neurosymbolic AI
Luca Andolfi, Eleonora Giunchiglia
Abstract
Neurosymbolic AI is growing in popularity thanks to its ability to combine neural perception and symbolic reasoning in end-to-end trainable models. However, recent findings reveal these are prone to shortcut reasoning, i.e., to learning unindented concepts-or neural predicates-which exploit spurious correlations to satisfy the symbolic constraints. In this paper, we address reasoning shortcuts at their root cause and we introduce Prototypical Neurosymbolic architectures. These models are able to satisfy the symbolic constraints (be right) because they have learnt the correct basic concepts (for the right reasons) and not because of spurious correlations, even in extremely low data regimes. Leveraging the theory of prototypical learning, we demonstrate that we can effectively avoid reasoning shortcuts by training the models to satisfy the background knowledge while taking into account the similarity of the input with respect to the handful of labelled datapoints. We extensively validate our approach on the recently proposed rsbench benchmark suite in a variety of settings and tasks with very scarce supervision: we show significant improvements in learning the right concepts both in synthetic tasks (MNIST-EvenOdd and Kand-Logic) and real-world, high-stake ones (BDD-OIA). Our findings pave the way to prototype grounding as an effective, annotation-efficient strategy for safe and reliable neurosymbolic learning. 1 Example 2 (Ex. 1, Cont'd.) Consider a restriction of the dataset where the ground truth for all datapoints is either g = (0, 6) or g ′ = (2, 8). Assume p θ (C | G) is a deterministic distribution so that p θ ((5, 5) | (2, 8)) ≈ 1.0 and p θ ((3, 3) | (0, 6)) ≈ 1.0. A NeSy predictor p θ whose (i) distribution p θ (C | G) is as shown before and (ii) likelihood L(p θ , D, K) is maximal, takes a RS.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers1
Ask how each one uses itBuilds on17
- Coherent Hierarchical Multi-Label Classification NetworksEleonora Giunchiglia, Thomas LukasiewiczNeurIPS 2020 · 142 citations
- Semantic Probabilistic Layers for Neuro-Symbolic LearningKareem Ahmed, Stefano Teso, Kai-Wei Chang, Guy Van den Broeck et al.NeurIPS 2022 · 133 citations
- Not All Neuro-Symbolic Concepts Are Created Equal: Analysis and Mitigation of Reasoning ShortcutsEmanuele Marconato, Stefano Teso, Antonio Vergari, Andrea PasseriniNeurIPS 2023 · 83 citations
- MultiplexNet: Towards Fully Satisfied Logical Constraints in Neural NetworksNick Hoernle, Rafael-Michael Karampatsis, Vaishak Belle, Kobi GalAAAI 2022 · 73 citations
- Interpretable Neural-Symbolic Concept ReasoningPietro Barbiero, Gabriele Ciravegna, Francesco Giannini, Mateo Espinosa Zarlenga et al.ICML 2023 · 68 citations
Related papers
- Mitigating Neuro-Symbolic Reasoning Shortcuts with Data-Driven Knowledge AugmentationYu-Feng Li, Xiao-Wen Yang, Wen-Da Wei, Jie-Jing Shao et al.KDD 2026
- A-NeSI: A Scalable Approximate Method for Probabilistic Neurosymbolic InferenceEmile van Krieken, Thiviyan Thanapalasingam, Jakub M. Tomczak, Frank van Harmelen et al.NeurIPS 2023 · 62 citations
- Analysis for Abductive Learning and Neural-Symbolic Reasoning ShortcutsXiaowen Yang, Wenda Wei, Jie-Jing Shao, Yufeng Li et al.ICML 2024 · 11 citations
- A learnability analysis on neuro-symbolic learningHao-Yuan He, Ming LiNeurIPS 2025 · 3 citations
- DeepProofLog: Efficient Proving in Deep Stochastic Logic ProgramsYing Jiao, Rodrigo Castellano Ontiveros, Luc De Raedt, Marco Gori et al.AAAI 2026
