Shift-Robust GNNs: Overcoming the Limitations of Localized Graph Training data
Qi Zhu, Natalia Ponomareva, Jiawei Han, Bryan Perozzi
摘要
There has been a recent surge of interest in designing Graph Neural Networks (GNNs) for semi-supervised learning tasks. Unfortunately this work has assumed that the nodes labeled for use in training were selected uniformly at random (i.e. are an IID sample). However in many real world scenarios gathering labels for graph nodes is both expensive and inherently biased -so this assumption can not be met. GNNs can suffer poor generalization when this occurs, by overfitting to superfluous regularities present in the training data. In this work we present a method, Shift-Robust GNN (SR-GNN), designed to account for distributional differences between biased training data and a graph's true inference distribution. SR-GNN adapts GNN models to the presence of distributional shift between the nodes labeled for training and the rest of the dataset. We illustrate the effectiveness of SR-GNN in a variety of experiments with biased training datasets on common GNN benchmark datasets for semi-supervised learning, where we see that SR-GNN outperforms other GNN baselines in accuracy, addressing at least ∼40% of the negative effects introduced by biased training data. On the largest dataset we consider, ogb-arxiv, we observe a 2% absolute improvement over the baseline and are able to mitigate 30% of the negative effects from training data bias 1 . Recently, GNNs have emerged as a way to combine graph structure with deep neural networks. Surprisingly, most work on semi-supervised learning using GNNs for node classification [15, 11, 1] have ignored this critical problem, and even the most recently proposed GNN benchmarks [12] assume that an independent and identically distributed (IID) sample is possible for training labels.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper51
- Handling Distribution Shifts on Graphs: An Invariance PerspectiveQitian Wu, Hengrui Zhang, Junchi Yan, David WipfICLR 2022 · 被引用 261 次
- Learning Invariant Graph Representations for Out-of-Distribution GeneralizationHaoyang Li, Ziwei Zhang, Xin Wang, Wenwu ZhuNeurIPS 2022 · 被引用 170 次
- Transfer Learning of Graph Neural Networks with Ego-graph Information MaximizationQi Zhu, Carl Yang, Yidan Xu, Haonan Wang 等NeurIPS 2021 · 被引用 140 次
- Dynamic Graph Neural Networks Under Spatio-Temporal Distribution ShiftZeyang Zhang, Xin Wang, Ziwei Zhang, Haoyang Li 等NeurIPS 2022 · 被引用 122 次
- Label-free Node Classification on Graphs with Large Language Models (LLMs)Zhikai Chen, Haitao Mao, Hongzhi Wen, Haoyu Han 等ICLR 2024 · 被引用 103 次
它引用的顶会 Paper5
- Open Graph Benchmark: Datasets for Machine Learning on GraphsWeihua Hu, Matthias Fey, Marinka Zitnik, Yuxiao Dong 等NeurIPS 2020 · 被引用 3,935 次
- Strategies for Pre-training Graph Neural NetworksWeihua Hu, Bowen Liu, Joseph Gomes, Marinka Zitnik 等ICLR 2020 · 被引用 1,744 次
- Unsupervised Domain Adaptive Graph Convolutional NetworksMan Wu, Shirui Pan, Chuan Zhou, Xiaojun Chang 等WWW 2020 · 被引用 221 次
- Continuous Graph Neural NetworksLouis-Pascal A. C. Xhonneux, Meng Qu, Jian TangICML 2020 · 被引用 194 次
- Transfer Learning of Graph Neural Networks with Ego-graph Information MaximizationQi Zhu, Carl Yang, Yidan Xu, Haonan Wang 等NeurIPS 2021 · 被引用 140 次
相关 Paper
- BA-GNN: On Learning Bias-Aware Graph Neural NetworkZhengyu Chen, Teng Xiao, Kun KuangICDE 2022 · 被引用 28 次
- NRGNN: Learning a Label Noise Resistant Graph Neural Network on Sparsely and Noisily Labeled GraphsEnyan Dai, Charu Aggarwal, Suhang WangKDD 2021 · 被引用 80 次
- Graph Out-of-Distribution Generalization via Causal InterventionQitian Wu, Fan Nie, Chenxiao Yang, Tianyi Bao 等WWW 2024 · 被引用 58 次
- Structural Fairness-aware Active Learning for Graph Neural NetworksHaoyu Han, Xiaorui Liu, Li Ma, MohamadAli Torkamani 等ICLR 2024 · 被引用 5 次
- Is Homophily a Necessity for Graph Neural Networks?Yao Ma, Xiaorui Liu, Neil Shah, Jiliang TangICLR 2022 · 被引用 295 次
