When to Pre-Train Graph Neural Networks? From Data Generation Perspective!
Yuxuan Cao, Jiarong Xu, Carl Yang, Jiaan Wang, Yunchao Zhang, Chunping Wang, Lei Chen, Yang Yang
Abstract
In recent years, graph pre-training has gained significant attention, focusing on acquiring transferable knowledge from unlabeled graph data to improve downstream performance. Despite these recent endeavors, the problem of negative transfer remains a major concern when utilizing graph pre-trained models to downstream tasks. Previous studies made great efforts on the issue of what to pre-train and how to pre-train by designing a variety of graph pre-training and fine-tuning strategies. However, there are cases where even the most advanced "pre-train and fine-tune" paradigms fail to yield distinct benefits. This paper introduces a generic framework W2PGNN to answer the crucial question of when to pre-train (.e., in what situations could we take advantage of graph pre-training) before performing effortful pre-training or fine-tuning. We start from a new perspective to explore the complex generative mechanisms from the pre-training data to downstream data. In particular, W2PGNN first fits the pre-training data into graphon bases, each element of graphon basis (i.e., a graphon) identifies a fundamental transferable pattern shared by a collection of pre-training graphs. All convex combinations of graphon bases give rise to a generator space, from which graphs generated form the solution space for those downstream data that can benefit from pre-training. In this manner, the feasibility of pre-training can be quantified as the generation probability of the downstream data from any generator in the generator space. W2PGNN offers three broad applications: providing the application scope of graph pre-trained models, quantifying the feasibility of pre-training, and assistance in selecting pre-training data to enhance downstream performance. We provide a theoretically sound solution for the first application and extensive empirical justifications for the latter two applications.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 55a610f6-57e8-4690-9faf-dd68ee23678dCited by top-tier papers19
- GFT: Graph Foundation Model with Transferable Tree VocabularyZehong Wang, Zheyuan Zhang, Nitesh V. Chawla, Chuxu Zhang et al.NeurIPS 2024 · 108 citations
- GraphKeeper: Graph Domain-Incremental Learning via Knowledge Disentanglement and PreservationZihao Guo, Qingyun Sun, Ziwei Zhang, Haonan Yuan et al.NeurIPS 2025 · 10 citations
- Cross-Domain Graph Data Scaling: A Showcase with Diffusion ModelsWenzhuo Tang, Haitao Mao, Danial Dervovic, Ivan Brugere et al.NeurIPS 2025 · 8 citations
- Adaptive Graph Integration for Cross-Domain Recommendation via Heterogeneous Graph CoordinatorsHengyu Zhang, Chunxu Shen, Xiangguo Sun, Jie Tan et al.SIGIR 2025 · 6 citations
- Improving the Robustness of Knowledge-Grounded Dialogue via Contrastive LearningJiaan Wang, Jianfeng Qu, Kexin Wang, Zhixu Li et al.AAAI 2024 · 5 citations
Builds on16
- Graph Contrastive Learning with AugmentationsYuning You, Tianlong Chen, Yongduo Sui, Ting Chen et al.NeurIPS 2020 · 3,042 citations
- Strategies for Pre-training Graph Neural NetworksWeihua Hu, Bowen Liu, Joseph Gomes, Marinka Zitnik et al.ICLR 2020 · 1,744 citations
- Contrastive Multi-View Representation Learning on GraphsKaveh Hassani, Amir Hosein Khas AhmadiICML 2020 · 1,663 citations
- InfoGraph: Unsupervised and Semi-supervised Graph-Level Representation Learning via Mutual Information MaximizationFan-Yun Sun, Jordan Hoffmann, Vikas Verma, Jian TangICLR 2020 · 1,010 citations
- GCC: Graph Contrastive Coding for Graph Neural Network Pre-TrainingJiezhong Qiu, Qibin Chen, Yuxiao Dong, Jing Zhang et al.KDD 2020 · 755 citations
Related papers
- Fine-Tuning Graph Neural Networks by Preserving Graph Generative PatternsYifei Sun, Qi Zhu, Yang Yang, Chunping Wang et al.AAAI 2024 · 21 citations
- Search to Fine-Tune Pre-Trained Graph Neural Networks for Graph-Level TasksZhili Wang, Shimin Di, Lei Chen, Xiaofang ZhouICDE 2024 · 6 citations
- Better with Less: A Data-Active Perspective on Pre-Training Graph Neural NetworksJiarong Xu, Renhong Huang, Xin Jiang, Yuxuan Cao et al.NeurIPS 2023 · 26 citations
- Measuring Task Similarity and Its Implication in Fine-Tuning Graph Neural NetworksRenhong Huang, Jiarong Xu, Xin Jiang, Chenglu Pan et al.AAAI 2024 · 14 citations
- GPPT: Graph Pre-training and Prompt Tuning to Generalize Graph Neural NetworksMingchen Sun, Kaixiong Zhou, Xin He, Ying Wang et al.KDD 2022 · 141 citations
