Evaluating Attribution for Graph Neural Networks
Benjamín Sánchez-Lengeling, Jennifer N. Wei, Brian K. Lee, Emily Reif, Peter Wang, Wesley Wei Qian, Kevin McCloskey, Lucy J. Colwell, Alexander B. Wiltschko
Abstract
Interpretability of machine learning models is critical to scientific understanding, AI safety, and debugging. Attribution is one approach to interpretability, which highlights input dimensions that are influential to a neural network's prediction. Evaluation of these methods is largely qualitative for image and text models, because acquiring ground truth attributions requires expensive and unreliable human judgment. Attribution has been comparatively understudied for graph neural networks (GNNs), a model class of growing importance that makes predictions on arbitrarily-sized graphs. Graph-valued data offer an opportunity to quantitatively benchmark attribution methods, because challenging synthetic graph problems have computable ground-truth attributions. In this work we adapt commonly-used attribution methods for GNNs and quantitatively evaluate them using the axes of attribution accuracy, stability, faithfulness and consistency. We make concrete recommendations for which attribution methods to use, and provide the data and code for our benchmarking suite. Rigorous and open source benchmarking of attribution methods in graphs could enable new methods development and broader use of attribution in real-world ML tasks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3e1f8458-be02-4f1b-9b1a-9b927df8d6e0Cited by top-tier papers16
- Evaluating Post-hoc Explanations for Graph Neural Networks via Robustness AnalysisJunfeng Fang, Wei Liu, Yuan Gao, Zemin Liu et al.NeurIPS 2023 · 39 citations
- Improving Molecular Graph Neural Network Explainability with Orthonormalization and Induced SparsityRyan Henderson, Djork-Arné Clevert, Floriane MontanariICML 2021 · 36 citations
- Composite Feature Selection Using Deep EnsemblesFergus Imrie, Alexander Norcliffe, Pietro Lió, Mihaela van der SchaarNeurIPS 2022 · 18 citations
- Explaining Graph Neural Networks via Structure-aware Interaction IndexNgoc Bui, Hieu Trung Nguyen, Viet Anh Nguyen, Rex YingICML 2024 · 16 citations
- GraphChef: Decision-Tree Recipes to Explain Graph Neural NetworksPeter Müller, Lukas Faber, Karolis Martinkus, Roger WattenhoferICLR 2024 · 11 citations
Builds on2
Related papers
- Reconsidering Faithfulness in Regular, Self-Explainable and Domain Invariant GNNsSteve Azzolin, Antonio Longa, Stefano Teso, Andrea PasseriniICLR 2025
- Faithful and Accurate Self-Attention Attribution for Message Passing Neural Networks via the Computation Tree ViewpointYong-Min Shin, Siqing Li, Xin Cao, Won-Yong ShinAAAI 2025 · 6 citations
- Multi-scale Explainer for Graph Neural NetworksLutong Wu, Shiying Cheng, Zhiqiang Wang, Jianqing Liang et al.ICML 2026
- Interpreting Graph Neural Networks for NLP With Differentiable Edge MaskingMichael Sejr Schlichtkrull, Nicola De Cao, Ivan TitovICLR 2021 · 287 citations
- GOAt: Explaining Graph Neural Networks via Graph Output AttributionShengyao Lu, Keith G. Mills, Jiao He, Bang Liu et al.ICLR 2024 · 16 citations
