Graph Inverse Style Transfer for Counterfactual Explainability
Bardh Prenkaj, Efstratios Zaradoukas, Gjergji Kasneci
Abstract
Counterfactual explainability seeks to uncover model decisions by identifying minimal changes to the input that alter the predicted outcome. This task becomes particularly challenging for graph data due to preserving structural integrity and semantic meaning. Unlike prior approaches that rely on forward perturbation mechanisms, we introduce Graph Inverse Style Transfer (GIST), the first framework to re-imagine graph counterfactual generation as a backtracking process, leveraging spectral style transfer. By aligning the global structure with the original input spectrum and preserving local content faithfulness, GIST produces valid counterfactuals as interpolations between the input style and counterfactual content. Tested on 8 binary and multi-class graph classification benchmarks, GIST achieves a remarkable +7.6% improvement in the validity of produced counterfactuals and significant gains (+45.5%) in faithfully explaining the true class distribution. Additionally, GIST's backtracking mechanism effectively mitigates overshooting the underlying predictor's decision boundary, minimizing the spectral differences between the input and the counterfactuals. These results challenge traditional forward perturbation methods, offering a novel perspective that advances graph explainability.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on5
- Learning and Evaluating Graph Neural Network Explanations based on Counterfactual and Factual ReasoningJuntao Tan, Shijie Geng, Zuohui Fu, Yingqiang Ge et al.WWW 2022 · 151 citations
- Robust Counterfactual Explanations on Graph Neural NetworksMohit Bajaj, Lingyang Chu, Zi Yu Xue, Jian Pei et al.NeurIPS 2021 · 140 citations
- Multimodal Motion Conditioned Diffusion Model for Skeleton-based Video Anomaly DetectionAlessandro Flaborea, Luca Collorone, Guido Maria D'Amely di Melendugno, Stefano D'Arrigo et al.ICCV 2023 · 83 citations
- CLEAR: Generative Counterfactual Explanations on GraphsJing Ma, Ruocheng Guo, Saumitra Mishra, Aidong Zhang et al.NeurIPS 2022 · 83 citations
- Unifying Evolution, Explanation, and Discernment: A Generative Approach for Dynamic Graph CounterfactualsBardh Prenkaj, Mario Villaizán-Vallelado, Tobias Leemann, Gjergji KasneciKDD 2024 · 1 citation
Related papers
- Generating In-Distribution Counterfactual Explanation for Graph Neural NetworksLinmao Chen, Chaobo He, Junwei Cheng, Chunying Li et al.AAAI 2026
- ATEX-CF: Attack-Informed Counterfactual Explanations for Graph Neural NetworksYu Zhang, Sean Bin Yang, Arijit Khan, Cuneyt Gurcan AkcoraICLR 2026 · 4 citations
- D4Explainer: In-distribution Explanations of Graph Neural Network via Discrete Denoising DiffusionJialin Chen, Shirley Wu, Abhijit Gupta, Rex YingNeurIPS 2023 · 31 citations
- Counterfactual Analysis on Large GraphsHsi-Wen Chen, Jian Pei, De-Nian Yang, Ming-Syan ChenKDD 2026
- UNR-Explainer: Counterfactual Explanations for Unsupervised Node Representation Learning ModelsHyunju Kang, Geonhee Han, Hogun ParkICLR 2024 · 8 citations
