Entropy-Based Logic Explanations of Neural Networks
Pietro Barbiero, Gabriele Ciravegna, Francesco Giannini, Pietro Lió, Marco Gori, Stefano Melacci
摘要
Explainable artificial intelligence has rapidly emerged since lawmakers have started requiring interpretable models for safety-critical domains. Concept-based neural networks have arisen as explainable-by-design methods as they leverage human-understandable symbols (i.e. concepts) to predict class memberships. However, most of these approaches focus on the identification of the most relevant concepts but do not provide concise, formal explanations of how such concepts are leveraged by the classifier to make predictions. In this paper, we propose a novel end-to-end differentiable approach enabling the extraction of logic explanations from neural networks using the formalism of First-Order Logic. The method relies on an entropy-based criterion which automatically identifies the most relevant concepts. We consider four different case studies to demonstrate that: (i) this entropy-based criterion enables the distillation of concise logic explanations in safety-critical domains from clinical data to computer vision; (ii) the proposed approach outperforms state-of-the-art white-box models in terms of classification accuracy and matches black box performances.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper17
- Concept Activation Regions: A Generalized Framework For Concept-Based ExplanationsJonathan Crabbé, Mihaela van der SchaarNeurIPS 2022 · 被引用 88 次
- Interpretable Neural-Symbolic Concept ReasoningPietro Barbiero, Gabriele Ciravegna, Francesco Giannini, Mateo Espinosa Zarlenga 等ICML 2023 · 被引用 68 次
- Global Concept-Based Interpretability for Graph Neural Networks via Neuron AnalysisHan Xuanyuan, Pietro Barbiero, Dobrik Georgiev, Lucie Charlotte Magister 等AAAI 2023 · 被引用 62 次
- Energy-Based Concept Bottleneck Models: Unifying Prediction, Concept Intervention, and Probabilistic InterpretationsXinyue Xu, Yi Qin, Lu Mi, Hao Wang 等ICLR 2024 · 被引用 32 次
- Dividing and Conquering a BlackBox to a Mixture of Interpretable Models: Route, Interpret, RepeatShantanu Ghosh, Ke Yu, Forough Arabshahi, Kayhan BatmanghelichICML 2023 · 被引用 15 次
它引用的顶会 Paper3
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- A Constraint-Based Approach to Learning and ExplanationGabriele Ciravegna, Francesco Giannini, Stefano Melacci, Marco Maggini 等AAAI 2020 · 被引用 16 次
- Self-Training With Noisy Student Improves ImageNet ClassificationQizhe Xie, Minh-Thang Luong, Eduard H. Hovy, Quoc V. LeCVPR 2020
相关 Paper
- A Peek Into the Reasoning of Neural Networks: Interpreting With Structural Visual ConceptsYunhao Ge, Yao Xiao, Zhi Xu, Meng Zheng 等CVPR 2021
- SIC: Similarity-Based Interpretable Image Classification with Neural NetworksTom Nuno Wolf, Emre Kavak, Fabian Bongratz, Christian WachingerICCV 2025 · 被引用 1 次
- Formal Abductive Latent Explanations for Prototype-Based NetworksJules Soria, Zakaria Chihani, Julien Girard-Satabin, Alban Grastien 等AAAI 2026 · 被引用 2 次
- DeXAR: Deep Explainable Sensor-Based Activity Recognition in Smart-Home EnvironmentsLuca Arrotta, Gabriele Civitarese, Claudio BettiniUbiComp 2022 · 被引用 40 次
- Learn to Explain Efficiently via Neural Logic Inductive LearningYuan Yang, Le SongICLR 2020 · 被引用 83 次
