Sufficient Reasons for Classifier Decisions in the Presence of Domain Constraints
Niku Gorji, Sasha Rubin
摘要
Recent work has unveiled a theory for reasoning about the decisions made by binary classifiers: a classifier describes a Boolean function, and the reasons behind an instance being classified as positive are the prime-implicants of the function that are satisfied by the instance. One drawback of these works is that they do not explicitly treat scenarios where the underlying data is known to be constrained, e.g., certain combinations of features may not exist, may not be observable, or may be required to be disregarded. We propose a more general theory, also based on prime-implicants, tailored to taking constraints into account. The main idea is to view classifiers as describing partial Boolean functions that are undefined on instances that do not satisfy the constraints. We prove that this simple idea results in more parsimonious reasons. That is, not taking constraints into account (e.g., ignoring, or taking them as negative instances) results in reasons that are subsumed by reasons that do take constraints into account. We illustrate this improved succinctness on synthetic classifiers and classifiers learnt from real data.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Local vs. Global Interpretability: A Computational Complexity PerspectiveShahaf Bassan, Guy Amir, Guy KatzICML 2024 · 被引用 28 次
- Solving Explainability Queries with Quantification: The Case of Feature RelevancyXuanxiang Huang, Yacine Izza, João Marques-SilvaAAAI 2023 · 被引用 16 次
- Eliminating the Impossible, Whatever Remains Must Be True: On Extracting and Applying Background Knowledge in the Context of Formal ExplanationsJinqiang Yu, Alexey Ignatiev, Peter J. Stuckey, Nina Narodytska 等AAAI 2023
- What makes an Ensemble (Un) Interpretable?Shahaf Bassan, Guy Amir, Meirav Zehavi, Guy KatzICML 2025
相关 Paper
- Trading Complexity for Sparsity in Random Forest ExplanationsGilles Audemard, Steve Bellart, Louenas Bounia, Frédéric Koriche 等AAAI 2022 · 被引用 58 次
- On the Computation of Necessary and Sufficient ExplanationsAdnan Darwiche, Chunxi JiAAAI 2022 · 被引用 33 次
- Unsupervised Causal Binary Concepts Discovery with VAE for Black-Box Model ExplanationThien Q. Tran, Kazuto Fukuchi, Youhei Akimoto, Jun SakumaAAAI 2022 · 被引用 11 次
- Learning Fair Naive Bayes Classifiers by Discovering and Eliminating Discrimination PatternsYooJung Choi, Golnoosh Farnadi, Behrouz Babaki, Guy Van den BroeckAAAI 2020 · 被引用 31 次
- Label-Descriptive Patterns and Their Application to Characterizing Classification ErrorsMichael A. Hedderich, Jonas Fischer, Dietrich Klakow, Jilles VreekenICML 2022 · 被引用 14 次
