PWSHAP: A Path-Wise Explanation Model for Targeted Variables
Lucile Ter-Minassian, Oscar Clivio, Karla DiazOrdaz, Robin J. Evans, Christopher C. Holmes
Abstract
Predictive black-box models can exhibit high accuracy but their opaque nature hinders their uptake in safety-critical deployment environments. Explanation methods (XAI) can provide confidence for decision-making through increased transparency. However, existing XAI methods are not tailored towards models in sensitive domains where one predictor is of special interest, such as a treatment effect in a clinical model, or ethnicity in policy models. We introduce Path-Wise Shapley effects (PWSHAP), a framework for assessing the targeted effect of a binary (e.g. treatment) variable from a complex outcome model. Our approach augments the predictive model with a user-defined directed acyclic graph (DAG). The method then uses the graph alongside on-manifold Shapley values to identify effects along causal pathways whilst maintaining robustness to adversarial attacks. We establish error bounds for the identified path-wise Shapley effects and for Shapley values. We show PWSHAP can perform local bias and mediation analyses with faithfulness to the model. Further, if the targeted variable is randomised we can quantify local effect modification. We demonstrate the resolution, interpretability, and true locality of our approach on examples and a real-world experiment.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers1
Ask how each one uses itBuilds on6
- The Many Shapley Values for Model ExplanationMukund Sundararajan, Amir NajmiICML 2020 · 799 citations
- Asymmetric Shapley values: incorporating causal knowledge into model-agnostic explainabilityChristopher Frye, Colin Rowat, Ilya FeigeNeurIPS 2020 · 246 citations
- Causal Shapley Values: Exploiting Causal Knowledge to Explain Individual Predictions of Complex ModelsTom Heskes, Evi Sijben, Ioan Gabriel Bucur, Tom ClaassenNeurIPS 2020 · 235 citations
- Sense and Sensitivity Analysis: Simple Post-Hoc Analysis of Bias Due to Unobserved ConfoundingVictor Veitch, Anisha ZaveriNeurIPS 2020 · 67 citations
- Explaining Algorithmic Fairness Through Fairness-Aware Causal Path DecompositionWeishen Pan, Sen Cui, Jiang Bian, Changshui Zhang et al.KDD 2021 · 29 citations
Related papers
- Beyond TreeSHAP: Efficient Computation of Any-Order Shapley Interactions for Tree EnsemblesMaximilian Muschalik, Fabian Fumagalli, Barbara Hammer, Eyke HüllermeierAAAI 2024 · 35 citations
- SHAP values via sparse Fourier representationAli Gorji, Andisheh Amrollahi, Andreas KrauseNeurIPS 2025 · 11 citations
- X-Hacking: The Threat of Misguided AutoMLRahul Sharma, Sumantrak Mukherjee, Andrea Sipka, Eyke Hüllermeier et al.ICML 2025
- Explanations of Black-Box Models based on Directional Feature InteractionsAria Masoomi, Davin Hill, Zhonghui Xu, Craig P. Hersh et al.ICLR 2022 · 26 citations
- Linear tree shapPeng Yu, Albert Bifet, Jesse Read, Chao XuNeurIPS 2022 · 27 citations
