Sufficient Reasons for Classifier Decisions in the Presence of Domain Constraints
Niku Gorji, Sasha Rubin
Abstract
Recent work has unveiled a theory for reasoning about the decisions made by binary classifiers: a classifier describes a Boolean function, and the reasons behind an instance being classified as positive are the prime-implicants of the function that are satisfied by the instance. One drawback of these works is that they do not explicitly treat scenarios where the underlying data is known to be constrained, e.g., certain combinations of features may not exist, may not be observable, or may be required to be disregarded. We propose a more general theory, also based on prime-implicants, tailored to taking constraints into account. The main idea is to view classifiers as describing partial Boolean functions that are undefined on instances that do not satisfy the constraints. We prove that this simple idea results in more parsimonious reasons. That is, not taking constraints into account (e.g., ignoring, or taking them as negative instances) results in reasons that are subsumed by reasons that do take constraints into account. We illustrate this improved succinctness on synthetic classifiers and classifiers learnt from real data.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e67a1a57-2b1c-4756-a26a-1e97756221e8Cited by top-tier papers4
- Local vs. Global Interpretability: A Computational Complexity PerspectiveShahaf Bassan, Guy Amir, Guy KatzICML 2024 · 28 citations
- Solving Explainability Queries with Quantification: The Case of Feature RelevancyXuanxiang Huang, Yacine Izza, João Marques-SilvaAAAI 2023 · 16 citations
- Eliminating the Impossible, Whatever Remains Must Be True: On Extracting and Applying Background Knowledge in the Context of Formal ExplanationsJinqiang Yu, Alexey Ignatiev, Peter J. Stuckey, Nina Narodytska et al.AAAI 2023
- What makes an Ensemble (Un) Interpretable?Shahaf Bassan, Guy Amir, Meirav Zehavi, Guy KatzICML 2025
Related papers
- Trading Complexity for Sparsity in Random Forest ExplanationsGilles Audemard, Steve Bellart, Louenas Bounia, Frédéric Koriche et al.AAAI 2022 · 58 citations
- On the Computation of Necessary and Sufficient ExplanationsAdnan Darwiche, Chunxi JiAAAI 2022 · 33 citations
- Unsupervised Causal Binary Concepts Discovery with VAE for Black-Box Model ExplanationThien Q. Tran, Kazuto Fukuchi, Youhei Akimoto, Jun SakumaAAAI 2022 · 11 citations
- Learning Fair Naive Bayes Classifiers by Discovering and Eliminating Discrimination PatternsYooJung Choi, Golnoosh Farnadi, Behrouz Babaki, Guy Van den BroeckAAAI 2020 · 31 citations
- Label-Descriptive Patterns and Their Application to Characterizing Classification ErrorsMichael A. Hedderich, Jonas Fischer, Dietrich Klakow, Jilles VreekenICML 2022 · 14 citations
