On Locality of Local Explanation Models
Sahra Ghalebikesabi, Lucile Ter-Minassian, Karla DiazOrdaz, Chris C. Holmes
Abstract
Shapley values provide model agnostic feature attributions for model outcome at a particular instance by simulating feature absence under a global population distribution. The use of a global population can lead to potentially misleading results when local model behaviour is of interest. Hence we consider the formulation of neighbourhood reference distributions that improve the local interpretability of Shapley values. By doing so, we find that the Nadaraya-Watson estimator, a well-studied kernel regressor, can be expressed as a self-normalised importance sampling estimator. Empirically, we observe that Neighbourhood Shapley values identify meaningful sparse feature relevance attributions that provide insight into local model behaviour, complimenting conventional Shapley analysis. They also increase on-manifold explainability and robustness to the construction of adversarial classifiers. * equal contribution Preprint. Under review.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 850d1799-2205-4a6b-99d7-1eeddec04ac7Cited by top-tier papers14
- RKHS-SHAP: Shapley Values for Kernel MethodsSiu Lun Chau, Robert Hu, Javier González, Dino SejdinovicNeurIPS 2022 · 49 citations
- Making Sense of Dependence: Efficient Black-box Explanations Using Dependence MeasurePaul Novello, Thomas Fel, David VigourouxNeurIPS 2022 · 48 citations
- Explaining the Uncertain: Stochastic Shapley Values for Gaussian Process ModelsSiu Lun Chau, Krikamol Muandet, Dino SejdinovicNeurIPS 2023 · 35 citations
- Robust Models Are More Interpretable Because Attributions Look NormalZifan Wang, Matt Fredrikson, Anupam DattaICML 2022 · 33 citations
- Unfooling Perturbation-Based Post Hoc ExplainersZachariah Carmichael, Walter J. ScheirerAAAI 2023 · 18 citations
Builds on3
- The Many Shapley Values for Model ExplanationMukund Sundararajan, Amir NajmiICML 2020 · 799 citations
- Understanding Global Feature Contributions With Additive Importance MeasuresIan Covert, Scott M. Lundberg, Su-In LeeNeurIPS 2020 · 476 citations
- Asymmetric Shapley values: incorporating causal knowledge into model-agnostic explainabilityChristopher Frye, Colin Rowat, Ilya FeigeNeurIPS 2020 · 246 citations
Related papers
- From global to local MDI variable importances for random forests and when they are Shapley valuesAntonio Sutera, Gilles Louppe, Vân Anh Huynh-Thu, Louis Wehenkel et al.NeurIPS 2021 · 16 citations
- Shapley explainability on the data manifoldChristopher Frye, Damien de Mijolla, Tom Begley, Laurence Cowton et al.ICLR 2021 · 125 citations
- Sample based Explanations via Generalized RepresentersChe-Ping Tsai, Chih-Kuan Yeh, Pradeep RavikumarNeurIPS 2023 · 13 citations
- Interventional SHAP Values and Interaction Values for Piecewise Linear Regression TreesArtjom Zern, Klaus Broelemann, Gjergji KasneciAAAI 2023 · 27 citations
- Linear tree shapPeng Yu, Albert Bifet, Jesse Read, Chao XuNeurIPS 2022 · 27 citations
