Learning Posterior Predictive Distributions for Node Classification from Synthetic Graph Priors
Jeongwhan Choi, Jongwoo Kim, Woosung Kang, Noseong Park
摘要
One of the most challenging problems in graph machine learning is generalizing across graphs with diverse properties. Graph neural networks (GNNs) face a fundamental limitation: they require separate training for each new graph, preventing universal generalization across diverse graph datasets. A critical challenge facing GNNs lies in their reliance on labeled training data for each individual graph, a requirement that hinders the capacity for universal node classification due to the heterogeneity inherent in graphs --- differences in homophily levels, community structures, and feature distributions across datasets. Inspired by the success of large language models (LLMs) that achieve in-context learning through massive-scale pre-training on diverse datasets, we introduce NodePFN. This universal node classification method generalizes to arbitrary graphs without graph-specific training. NodePFN learns posterior predictive distributions (PPDs) by training only on thousands of synthetic graphs generated from carefully designed priors. Our synthetic graph generation covers real-world graphs through the use of random networks with controllable homophily levels and structural causal models for complex feature-label relationships. We develop a dual-branch architecture combining context-query attention mechanisms with local message passing to enable graph-aware in-context learning. Extensive evaluation on 23 benchmarks demonstrates that a single pre-trained NodePFN achieves 71.27% average accuracy. These results validate that universal graph learning patterns can be effectively learned from synthetic priors, establishing a new paradigm for generalization in node classification.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- GraphPFN: A Prior-Data Fitted Graph Foundation ModelDmitry Eremeev, Oleg Platonov, Gleb Bazhenov, Artem Babenko 等ICML 2026 · 被引用 15 次
- Node4All: Learning Node Representation Beyond DatasetsDooho Lee, Jaemin YooKDD 2026 · 被引用 2 次
- One Sequential Recommendation Model Pretrained from Synthetic Priors Predicts Multiple DatasetsWoosung Kang, Jiwon Jeong, Jonghyeok Shin, Jeongwhan Choi 等KDD 2026
它引用的顶会 Paper22
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- LightGCN: Simplifying and Powering Graph Convolution Network for RecommendationXiangnan He, Kuan Deng, Xiang Wang, Yan Li 等SIGIR 2020 · 被引用 4,448 次
- Beyond Homophily in Graph Neural Networks: Current Limitations and Effective DesignsJiong Zhu, Yujun Yan, Lingxiao Zhao, Mark Heimann 等NeurIPS 2020 · 被引用 1,490 次
- Geom-GCN: Geometric Graph Convolutional NetworksHongbin Pei, Bingzhe Wei, Kevin Chen-Chuan Chang, Yu Lei 等ICLR 2020 · 被引用 1,445 次
- Beyond Low-frequency Information in Graph Convolutional NetworksDeyu Bo, Xiao Wang, Chuan Shi, Huawei ShenAAAI 2021 · 被引用 773 次
相关 Paper
- Non-Homophilic Graph Pre-Training and Prompt LearningXingtong Yu, Jie Zhang, Yuan Fang, Renhe JiangKDD 2025 · 被引用 6 次
- Adaptive Universal Generalized PageRank Graph Neural NetworkEli Chien, Jianhao Peng, Pan Li, Olgica MilenkovicICLR 2021 · 被引用 93 次
- Strategies for Pre-training Graph Neural NetworksWeihua Hu, Bowen Liu, Joseph Gomes, Marinka Zitnik 等ICLR 2020 · 被引用 1,744 次
- HetGPT: Harnessing the Power of Prompt Tuning in Pre-Trained Heterogeneous Graph Neural NetworksYihong Ma, Ning Yan, Jiayu Li, Masood S. Mortazavi 等WWW 2024 · 被引用 50 次
- GraphUniverse: Synthetic Graph Generation for Evaluating Inductive GeneralizationLouis Van Langendonck, Guillermo Bernardez, Nina Miolane, Pere Barlet-RosICLR 2026 · 被引用 1 次
