Shortcuts and Identifiability in Concept-based Models from a Neuro-Symbolic Lens
Samuele Bortolotti, Emanuele Marconato, Paolo Morettin, Andrea Passerini, Stefano Teso
Abstract
Concept-based Models are neural networks that learn a concept extractor to map inputs to high-level concepts and an inference layer to translate these into predictions. Ensuring these modules produce interpretable concepts and behave reliably in out-of-distribution is crucial, yet the conditions for achieving this remain unclear. We study this problem by establishing a novel connection between Concept-based Models and reasoning shortcuts (RSs), a common issue where models achieve high accuracy by learning low-quality concepts, even when the inference layer is fixed and provided upfront. Specifically, we extend RSs to the more complex setting of Concept-based Models and derive theoretical conditions for identifying both the concepts and the inference layer. Our empirical results highlight the impact of RSs and show that existing methods, even combined with multiple natural mitigation strategies, often fail to meet these conditions in practice.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c6186428-da04-4c22-bf90-eb37236d08feCited by top-tier papers5
- Object-Centric Concept-BottlenecksDavid Steinmann, Wolfgang Stammer, Antonia Wüst, Kristian KerstingNeurIPS 2025 · 12 citations
- Deferring Concept Bottleneck Models: Learning to Defer Interventions to Inaccurate ExpertsAndrea Pugnana, Riccardo Massidda, Francesco Giannini, Pietro Barbiero et al.NeurIPS 2025 · 11 citations
- When Does Closeness in Distribution Imply Representational Similarity? An Identifiability PerspectiveBeatrix M. G. Nielsen, Emanuele Marconato, Andrea Dittadi, Luigi GreseleNeurIPS 2025 · 7 citations
- GNN Explanations that do not Explain and How to find ThemSteve Azzolin, Stefano Teso, Bruno Lepri, Andrea Passerini et al.ICLR 2026 · 4 citations
- Neural Concept Verifier: Scaling Prover-Verifier Games via Concept EncodingsBerkant Turan, Suhrab Asadulla, David Steinmann, Kristian Kersting et al.ICML 2026
Builds on41
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Concept Bottleneck ModelsPang Wei Koh, Thao Nguyen, Yew Siang Tang, Stephen Mussmann et al.ICML 2020 · 1,233 citations
- Self-Supervised Learning with Data Augmentations Provably Isolates Content from StyleJulius von Kügelgen, Yash Sharma, Luigi Gresele, Wieland Brendel et al.NeurIPS 2021 · 421 citations
- Addressing Leakage in Concept Bottleneck ModelsMarton Havasi, Sonali Parbhoo, Finale Doshi-VelezNeurIPS 2022 · 163 citations
- Coherent Hierarchical Multi-Label Classification NetworksEleonora Giunchiglia, Thomas LukasiewiczNeurIPS 2020 · 142 citations
Related papers
- Not All Neuro-Symbolic Concepts Are Created Equal: Analysis and Mitigation of Reasoning ShortcutsEmanuele Marconato, Stefano Teso, Antonio Vergari, Andrea PasseriniNeurIPS 2023 · 83 citations
- Causally Reliable Concept Bottleneck ModelsGiovanni de Felice, Arianna Casanova Flores, Francesco De Santis, Silvia Santini et al.NeurIPS 2025 · 20 citations
- Understanding Inter-Concept Relationships in Concept-Based ModelsNaveen Raman, Mateo Espinosa Zarlenga, Mateja JamnikICML 2024 · 12 citations
- Analysis for Abductive Learning and Neural-Symbolic Reasoning ShortcutsXiaowen Yang, Wenda Wei, Jie-Jing Shao, Yufeng Li et al.ICML 2024 · 11 citations
- Mitigating Neuro-Symbolic Reasoning Shortcuts with Data-Driven Knowledge AugmentationYu-Feng Li, Xiao-Wen Yang, Wen-Da Wei, Jie-Jing Shao et al.KDD 2026
