CLEAR: Generative Counterfactual Explanations on Graphs
Jing Ma, Ruocheng Guo, Saumitra Mishra, Aidong Zhang, Jundong Li
Abstract
Counterfactual explanations promote explainability in machine learning models by answering the question "how should an input instance be perturbed to obtain a desired predicted label?". The comparison of this instance before and after perturbation can enhance human interpretation. Most existing studies on counterfactual explanations are limited in tabular data or image data. In this work, we study the problem of counterfactual explanation generation on graphs. A few studies have explored counterfactual explanations on graphs, but many challenges of this problem are still not well-addressed: 1) optimizing in the discrete and disorganized space of graphs; 2) generalizing on unseen graphs; and 3) maintaining the causality in the generated counterfactuals without prior knowledge of the causal model. To tackle these challenges, we propose a novel framework CLEAR which aims to generate counterfactual explanations on graphs for graph-level prediction models. Specifically, CLEAR leverages a graph variational autoencoder based mechanism to facilitate its optimization and generalization, and promotes causality by leveraging an auxiliary variable to better identify the underlying causal model. Extensive experiments on both synthetic and real-world graphs validate the superiority of CLEAR over the state-of-the-art methods in different aspects. 36th Conference on Neural Information Processing Systems (NeurIPS 2022).
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 56a9eef1-edf2-41d3-af5e-d335df7cbd5fCited by top-tier papers20
- D4Explainer: In-distribution Explanations of Graph Neural Network via Discrete Denoising DiffusionJialin Chen, Shirley Wu, Abhijit Gupta, Rex YingNeurIPS 2023 · 31 citations
- GraphTrail: Translating GNN Predictions into Human-Interpretable Logical RulesBurouj Armgaan, Manthan Dalmia, Sourav Medya, Sayan RanuNeurIPS 2024 · 28 citations
- GNNX-BENCH: Unravelling the Utility of Perturbation-based GNN Explainers through In-depth BenchmarkingMert Kosan, Samidha Verma, Burouj Armgaan, Khushbu Pahwa et al.ICLR 2024 · 21 citations
- Graph Neural Networks for Vulnerability Detection: A Counterfactual ExplanationZhaoyang Chu, Yao Wan, Qian Li, Yang Wu et al.ISSTA 2024 · 19 citations
- Towards Robust Fidelity for Evaluating Explainability of Graph Neural NetworksXu Zheng, Farhad Shirani, Tianchun Wang, Wei Cheng et al.ICLR 2024 · 17 citations
Builds on2
- Algorithmic recourse under imperfect causal knowledge: a probabilistic approachAmir-Hossein Karimi, Bodo Julius von Kügelgen, Bernhard Schölkopf, Isabel ValeraNeurIPS 2020 · 224 citations
- Robust Counterfactual Explanations on Graph Neural NetworksMohit Bajaj, Lingyang Chu, Zi Yu Xue, Jian Pei et al.NeurIPS 2021 · 140 citations
Related papers
- Generating In-Distribution Counterfactual Explanation for Graph Neural NetworksLinmao Chen, Chaobo He, Junwei Cheng, Chunying Li et al.AAAI 2026
- ATEX-CF: Attack-Informed Counterfactual Explanations for Graph Neural NetworksYu Zhang, Sean Bin Yang, Arijit Khan, Cuneyt Gurcan AkcoraICLR 2026 · 4 citations
- Graph Inverse Style Transfer for Counterfactual ExplainabilityBardh Prenkaj, Efstratios Zaradoukas, Gjergji KasneciICML 2025
- UNR-Explainer: Counterfactual Explanations for Unsupervised Node Representation Learning ModelsHyunju Kang, Geonhee Han, Hogun ParkICLR 2024 · 8 citations
- CausalVAE: Disentangled Representation Learning via Neural Structural Causal ModelsMengyue Yang, Furui Liu, Zhitang Chen, Xinwei Shen et al.CVPR 2021
