Formal Abductive Latent Explanations for Prototype-Based Networks
Jules Soria, Zakaria Chihani, Julien Girard-Satabin, Alban Grastien, Romain Xu-Darme, Daniela Cancila
摘要
Case-based reasoning networks are machine-learning models that make predictions based on similarity between the input and prototypical parts of training samples, called prototypes. Such models are able to explain each decision by pointing to the prototypes that contributed the most to the final outcome. As the explanation is a core part of the prediction, they are often qualified as "interpretable by design". While promising, we show that such explanations are sometimes misleading, which hampers their usefulness in safety-critical contexts. In particular, several instances may lead to different predictions and yet have the same explanation. Drawing inspiration from the field of formal eXplainable AI (FXAI), we propose Abductive Latent Explanations (ALEs), a formalism to express sufficient conditions on the intermediate (latent) representation of the instance that imply the prediction. Our approach combines the inherent interpretability of case-based reasoning models and the guarantees provided by formal XAI. We propose a solver-free and scalable algorithm for generating ALEs based on three distinct paradigms, compare them, and present the feasibility of our approach on diverse datasets for both standard and fine-grained image classification. The associated code can be found here 1 .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Formal Mechanistic Interpretability: Automated Circuit Discovery with Provable GuaranteesItamar Hadad, Guy Katz, Shahaf BassanICLR 2026 · 被引用 10 次
- Provably Explaining Neural Additive ModelsShahaf Bassan, Yizhak Yisrael Elboher, Tobias Ladner, Volkan Şahin 等ICLR 2026 · 被引用 3 次
- Verified SHAP: Provable Bounds for Exact Shapley Values of Neural NetworksDavid Boetius, Shahaf Bassan, Guy Katz, Stefan Leue 等ICML 2026
它引用的顶会 Paper8
- Concept Bottleneck ModelsPang Wei Koh, Thao Nguyen, Yew Siang Tang, Stephen Mussmann 等ICML 2020 · 被引用 1,233 次
- What I Cannot Predict, I Do Not Understand: A Human-Centered Evaluation Framework for Explainability MethodsJulien Colin, Thomas Fel, Rémi Cadène, Thomas SerreNeurIPS 2022 · 被引用 147 次
- ProtoPShare: Prototypical Parts Sharing for Similarity Discovery in Interpretable Image ClassificationDawid Rymarczyk, Lukasz Struski, Jacek Tabor, Bartosz ZielinskiKDD 2021 · 被引用 78 次
- VeriX: Towards Verified Explainability of Deep Neural NetworksMin Wu, Haoze Wu, Clark W. BarrettNeurIPS 2023 · 被引用 39 次
- CRAFT: Concept Recursive Activation FacTorization for ExplainabilityThomas Fel, Agustin Martin Picard, Louis Béthune, Thibaut Boissin 等CVPR 2023
相关 Paper
- Keep the Faith: Faithful Explanations in Convolutional Neural Networks for Case-Based ReasoningTom Nuno Wolf, Fabian Bongratz, Anne-Marie Rickmann, Sebastian Pölsterl 等AAAI 2024 · 被引用 11 次
- Entropy-Based Logic Explanations of Neural NetworksPietro Barbiero, Gabriele Ciravegna, Francesco Giannini, Pietro Lió 等AAAI 2022 · 被引用 97 次
- Deformable ProtoPNet: An Interpretable Image Classifier Using Deformable PrototypesJon Donnelly, Alina Jade Barnett, Chaofan ChenCVPR 2022 · 被引用 101 次
- Learning to Select Prototypical Parts for Interpretable Sequential Data ModelingYifei Zhang, Neng Gao, Cunqing MaAAAI 2023 · 被引用 10 次
- SIC: Similarity-Based Interpretable Image Classification with Neural NetworksTom Nuno Wolf, Emre Kavak, Fabian Bongratz, Christian WachingerICCV 2025 · 被引用 1 次
