Online GNN Evaluation Under Test-time Graph Distribution Shifts
Xin Zheng, Dongjin Song, Qingsong Wen, Bo Du, Shirui Pan
摘要
Evaluating the performance of a well-trained GNN model on real-world graphs is a pivotal step for reliable GNN online deployment and serving. Due to a lack of test node labels and unknown potential training-test graph data distribution shifts, conventional model evaluation encounters limitations in calculating performance metrics (e.g., test error) and measuring graph data-level discrepancies, particularly when the training graph used for developing GNNs remains unobserved during test time. In this paper, we study a new research problem, online GNN evaluation, which aims to provide valuable insights into the well-trained GNNs's ability to effectively generalize to real-world unlabeled graphs under the test-time graph distribution shifts. Concretely, we develop an effective learning behavior discrepancy score, dubbed LeBeD, to estimate the test-time generalization errors of well-trained GNN models. Through a novel GNN re-training strategy with a parameter-free optimality criterion, the proposed LeBeD comprehensively integrates learning behavior discrepancies from both node prediction and structure reconstruction perspectives. This enables the effective evaluation of the well-trained GNNs' ability to capture test node semantics and structural representations, making it an expressive metric for estimating the generalization error in online GNN evaluation. Extensive experiments on real-world test graphs under diverse graph distribution shifts could verify the effectiveness of the proposed method, revealing its strong correlation with ground-truth test errors on various well-trained GNN models.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Dynamic Graph Unlearning: A General and Efficient Post-Processing Method via Gradient TransformationHe Zhang, Bang Wu, Xiangwen Yang, Xingliang Yuan 等WWW 2025 · 被引用 16 次
- Shapley-Guided Utility Learning for Effective Graph Inference Data ValuationHongliang Chi, Qiong Wu, Zhengyi Zhou, Yao MaICLR 2025
- BiMark: Unbiased Multilayer Watermarking for Large Language ModelsXiaoyan Feng, He Zhang, Yanjun Zhang, Leo Yu Zhang 等ICML 2025
- N-ForGOT: Towards Not-forgetting and Generalization of Open Temporal Graph LearningLiping Wang, Xujia Li, Jingshu Peng, Yue Wang 等ICLR 2025
- GFMate: Empowering Graph Foundation Models with Test-time Prompt TuningYan Jiang, Ruihong Qiu, Zi HuangICML 2026
它引用的顶会 Paper23
- EvolveGCN: Evolving Graph Convolutional Networks for Dynamic GraphsAldo Pareja, Giacomo Domeniconi, Jie Chen, Tengfei Ma 等AAAI 2020 · 被引用 1,429 次
- Reasoning on Graphs: Faithful and Interpretable Large Language Model ReasoningLinhao Luo, Yuan-Fang Li, Gholamreza Haffari, Shirui PanICLR 2024 · 被引用 499 次
- Handling Distribution Shifts on Graphs: An Invariance PerspectiveQitian Wu, Hengrui Zhang, Junchi Yan, David WipfICLR 2022 · 被引用 261 次
- Unsupervised Domain Adaptive Graph Convolutional NetworksMan Wu, Shirui Pan, Chuan Zhou, Xiaojun Chang 等WWW 2020 · 被引用 221 次
- Leveraging unlabeled data to predict out-of-distribution performanceSaurabh Garg, Sivaraman Balakrishnan, Zachary Chase Lipton, Behnam Neyshabur 等ICLR 2022 · 被引用 160 次
相关 Paper
- GNNEvaluator: Evaluating GNN Performance On Unseen Graphs Without LabelsXin Zheng, Miao Zhang, Chunyang Chen, Soheila Molaei 等NeurIPS 2023 · 被引用 30 次
- Learning to Reweight for Generalizable Graph Neural NetworkZhengyu Chen, Teng Xiao, Kun Kuang, Zheqi Lv 等AAAI 2024 · 被引用 26 次
- Label Attentive Distillation for GNN-Based Graph ClassificationXiaobin Hong, Wenzhong Li, Chaoqun Wang, Mingkai Lin 等AAAI 2024 · 被引用 14 次
- Topology-Aware Dynamic Reweighting for Distribution Shifts on GraphWeihuang Zheng, Jiashuo Liu, Jiaxing Li, Jiayun Wu 等ICML 2025
- Demystifying Structural Disparity in Graph Neural Networks: Can One Size Fit All?Haitao Mao, Zhikai Chen, Wei Jin, Haoyu Han 等NeurIPS 2023 · 被引用 58 次
