Keep the Faith: Faithful Explanations in Convolutional Neural Networks for Case-Based Reasoning
Tom Nuno Wolf, Fabian Bongratz, Anne-Marie Rickmann, Sebastian Pölsterl, Christian Wachinger
摘要
Explaining predictions of black-box neural networks is crucial when applied to decision-critical tasks. Thus, attribution maps are commonly used to identify important image regions, despite prior work showing that humans prefer explanations based on similar examples. To this end, ProtoPNet learns a set of class-representative feature vectors (prototypes) for case-based reasoning. During inference, similarities of latent features to prototypes are linearly classified to form predictions and attribution maps are provided to explain the similarity. In this work, we evaluate whether architectures for case-based reasoning fulfill established axioms required for faithful explanations using the example of ProtoPNet. We show that such architectures allow the extraction of faithful explanations. However, we prove that the attribution maps used to explain the similarities violate the axioms. We propose a new procedure to extract explanations for trained ProtoPNets, named ProtoPFaith. Conceptually, these explanations are Shapley values, calculated on the similarity scores of each prototype. They allow to faithfully answer which prototypes are present in an unseen image and quantify each pixel’s contribution to that presence, thereby complying with all axioms. The theoretical violations of ProtoPNet manifest in our experiments on three datasets (CUB-200-2011, Stanford Dogs, RSNA) and five architectures (ConvNet, ResNet, ResNet50, WideResNet50, ResNeXt50). Our experiments show a qualitative difference between the explanations given by ProtoPNet and ProtoPFaith. Additionally, we quantify the explanations with the Area Over the Perturbation Curve, on which ProtoPFaith outperforms ProtoPNet on all experiments by a factor >10^3.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- A Robust Prototype-Based Network with Interpretable RBF Classifier FoundationsSascha Saralajew, Ashish Rana, Thomas Villmann, Ammar ShakerAAAI 2025 · 被引用 7 次
- Post-hoc Part-Prototype NetworksAndong Tan, Fengtao Zhou, Hao ChenICML 2024 · 被引用 7 次
- FaCT: Faithful Concept Traces for Explaining Neural Network DecisionsAmin Parchami-Araghi, Sukrut Rao, Jonas Fischer, Bernt SchieleNeurIPS 2025 · 被引用 1 次
- SIC: Similarity-Based Interpretable Image Classification with Neural NetworksTom Nuno Wolf, Emre Kavak, Fabian Bongratz, Christian WachingerICCV 2025 · 被引用 1 次
它引用的顶会 Paper11
- Understanding Global Feature Contributions With Additive Importance MeasuresIan Covert, Scott M. Lundberg, Su-In LeeNeurIPS 2020 · 被引用 476 次
- Sanity Checks for Saliency MetricsRichard Tomsett, Dan Harborne, Supriyo Chakraborty, Prudhvi Gurram 等AAAI 2020 · 被引用 204 次
- "Help Me Help the AI": Understanding How Explainability Can Support Human-AI InteractionSunnie S. Y. Kim, Elizabeth Anne Watkins, Olga Russakovsky, Ruth Fong 等CHI 2023 · 被引用 178 次
- How Can I Explain This to You? An Empirical Study of Deep Neural Network Explanation MethodsJeya Vikranth Jeyakumar, Joseph Noor, Yu-Hsi Cheng, Luis Garcia 等NeurIPS 2020 · 被引用 173 次
- The effectiveness of feature attribution methods and its correlation with automatic evaluation scoresGiang Nguyen, Daeyoung Kim, Anh NguyenNeurIPS 2021 · 被引用 128 次
相关 Paper
- This Looks Like Those: Illuminating Prototypical Concepts Using Multiple VisualizationsChiyu Ma, Brandon Zhao, Chaofan Chen, Cynthia RudinNeurIPS 2023 · 被引用 53 次
- Deformable ProtoPNet: An Interpretable Image Classifier Using Deformable PrototypesJon Donnelly, Alina Jade Barnett, Chaofan ChenCVPR 2022 · 被引用 101 次
- Interpretable Image Classification via Non-parametric Part Prototype LearningZhijie Zhu, Lei Fan, Maurice Pagnucco, Yang SongCVPR 2025
- This Looks Like It Rather Than That: ProtoKNN For Similarity-Based ClassifiersYuki Ukai, Tsubasa Hirakawa, Takayoshi Yamashita, Hironobu FujiyoshiICLR 2023
- Formal Abductive Latent Explanations for Prototype-Based NetworksJules Soria, Zakaria Chihani, Julien Girard-Satabin, Alban Grastien 等AAAI 2026 · 被引用 2 次
