Interpreting Graph Neural Networks for NLP With Differentiable Edge Masking
Michael Sejr Schlichtkrull, Nicola De Cao, Ivan Titov
Abstract
Graph neural networks (GNNs) have become a popular approach to integrating structural inductive biases into NLP models. However, there has been little work on interpreting them, and specifically on understanding which parts of the graphs (e.g. syntactic trees or co-reference structures) contribute to a prediction. In this work, we introduce a post-hoc method for interpreting the predictions of GNNs which identifies unnecessary edges. Given a trained GNN model, we learn a simple classifier that, for every edge in every layer, predicts if that edge can be dropped. We demonstrate that such a classifier can be trained in a fully differentiable fashion, employing stochastic gates and encouraging sparsity through the expected norm. We use our technique as an attribution method to analyze GNN models for two tasks -- question answering and semantic role labeling -- providing insights into the information flow in these models. We show that we can drop a large proportion of edges without deteriorating the performance of the model, while we can analyse the remaining edges for interpreting model predictions.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 95691ce1-414b-497d-a2ed-26cbd526e0c2Cited by top-tier papers62
- On Explainability of Graph Neural Networks via Subgraph ExplorationsHao Yuan, Haiyang Yu, Jie Wang, Kang Li et al.ICML 2021 · 498 citations
- Interpretable and Generalizable Graph Learning via Stochastic Attention MechanismSiqi Miao, Mia Liu, Pan LiICML 2022 · 288 citations
- Towards Multi-Grained Explainability for Graph Neural NetworksXiang Wang, Ying-Xin Wu, An Zhang, Xiangnan He et al.NeurIPS 2021 · 105 citations
- Extending the Nested Model for User-Centric XAI: A Design Study on GNN-based Drug RepurposingQianwen Wang, Kexin Huang, Payal Chandak, Marinka Zitnik et al.IEEE VIS 2022 · 83 citations
- Layer-refined Graph Convolutional Networks for RecommendationXin Zhou, Donghui Lin, Yong Liu, Chunyan MiaoICDE 2023 · 81 citations
Builds on4
- Parameterized Explainer for Graph Neural NetworkDongsheng Luo, Wei Cheng, Dongkuan Xu, Wenchao Yu et al.NeurIPS 2020 · 888 citations
- Restricting the Flow: Information Bottlenecks for AttributionKarl Schulz, Leon Sixt, Federico Tombari, Tim LandgrafICLR 2020 · 220 citations
- Towards Hierarchical Importance Attribution: Explaining Compositional Semantics for Neural Sequence ModelsXisen Jin, Zhongyu Wei, Junyi Du, Xiangyang Xue et al.ICLR 2020 · 55 citations
- How do Decisions Emerge across Layers in Neural Models? Interpretation with Differentiable MaskingNicola De Cao, Michael Sejr Schlichtkrull, Wilker Aziz, Ivan TitovEMNLP 2020 · 13 citations
Related papers
- Evaluating Attribution for Graph Neural NetworksBenjamín Sánchez-Lengeling, Jennifer N. Wei, Brian K. Lee, Emily Reif et al.NeurIPS 2020 · 159 citations
- Interpretable Sparsification of Brain Graphs: Better Practices and Effective Designs for Graph Neural NetworksGaotang Li, Marlena Duda, Xiang Zhang, Danai Koutra et al.KDD 2023 · 7 citations
- Adaptive Node Feature Selection for Graph Neural NetworksMadeline Navarro, Ali Azizpour, Santiago SegarraICML 2026
- Generative Causal Explanations for Graph Neural NetworksWanyu Lin, Hao Lan, Baochun LiICML 2021 · 217 citations
- Discovering Invariant Rationales for Graph Neural NetworksYingxin Wu, Xiang Wang, An Zhang, Xiangnan He et al.ICLR 2022 · 313 citations
