On Generating Plausible Counterfactual and Semi-Factual Explanations for Deep Learning
Eoin M. Kenny, Mark T. Keane
摘要
There is a growing concern that the recent progress made in AI, especially regarding the predictive competence of deep learning models, will be undermined by a failure to properly explain their operation and outputs. In response to this disquiet, counterfactual explanations have become massively popular in eXplainable AI (XAI) due to their proposed computational, psychological, and legal benefits. In contrast however, semi-factuals, which are a similar way humans commonly explain their reasoning, have surprisingly received no attention. Most counterfactual methods address tabular rather than image data, partly due to the latter's non-discrete nature making good counterfactuals difficult to define. Additionally, generating plausible looking explanations which lie on the data manifold is another issue which hampers progress. This paper advances a novel method for generating plausible counterfactuals (and semi-factuals) for black-box CNN classifiers doing computer vision. The present method, called PlausIble Exceptionality-based Contrastive Explanations (PIECE), modifies all "exceptional" features in a test image to be "normal" from the perspective of the counterfactual class (hence concretely defining a counterfactual). Two controlled experiments compare this method to others in the literature, showing that PIECE not only generates the most plausible counterfactuals on several measures, but also the best semi-factuals.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- Learning Support and Trivial Prototypes for Interpretable Image ClassificationChong Wang, Yuyuan Liu, Yuanhong Chen, Fengbei Liu 等ICCV 2023 · 被引用 50 次
- A Rationale-Centric Framework for Human-in-the-loop Machine LearningJinghui Lu, Linyi Yang, Brian MacNamee, Yue ZhangACL 2022 · 被引用 46 次
- Bayes-TrEx: a Bayesian Sampling Approach to Model Transparency by ExampleSerena Booth, Yilun Zhou, Ankit Shah, Julie ShahAAAI 2021 · 被引用 20 次
- GAM Coach: Towards Interactive and User-centered Algorithmic RecourseZijie J. Wang, Jennifer Wortman Vaughan, Rich Caruana, Duen Horng ChauCHI 2023 · 被引用 18 次
- The Utility of "Even if" Semifactual Explanation to Optimise Positive OutcomesEoin M. Kenny, Weipeng HuangNeurIPS 2023 · 被引用 16 次
它引用的顶会 Paper1
相关 Paper
- Towards More Faithful Natural Language Explanation Using Multi-Level Contrastive Learning in VQAChengen Lai, Shengli Song, Shiqi Meng, Jingyang Li 等AAAI 2024 · 被引用 12 次
- CoCoX: Generating Conceptual and Counterfactual Explanations via Fault-LinesArjun R. Akula, Shuai Wang, Song-Chun ZhuAAAI 2020 · 被引用 102 次
- Designing Counterfactual Generators using Deep Model InversionJayaraman J. Thiagarajan, Vivek Sivaraman Narayanaswamy, Deepta Rajan, Jason Liang 等NeurIPS 2021 · 被引用 25 次
- Accurate Explanation Model for Image Classifiers using Class Association EmbeddingRuitao Xie, Jingbang Chen, Limai Jiang, Rui Xiao 等ICDE 2024 · 被引用 12 次
- Looking in the Mirror: A Faithful Counterfactual Explanation Method for Interpreting Deep Image Classification ModelsTownim Faisal Chowdhury, Vu Minh Hieu Phan, Kewen Liao, Nanyu Dong 等ICCV 2025
