Pre-training on Large-Scale Heterogeneous Graph
Xunqiang Jiang, Tianrui Jia, Yuan Fang, Chuan Shi, Zhe Lin, Hui Wang
摘要
Graph neural networks (GNNs) emerge as the state-of-the-art representation learning methods on graphs and often rely on a large amount of labeled data to achieve satisfactory performance. Recently, in order to relieve the label scarcity issues, some works propose to pre-train GNNs in a self-supervised manner by distilling transferable knowledge from the unlabeled graph structures. Unfortunately, these pre-training frameworks mainly target at homogeneous graphs, while real interaction systems usually constitute large-scale heterogeneous graphs, containing different types of nodes and edges, which leads to new challenges on structure heterogeneity and scalability for graph pre-training. In this paper, we first study the problem of pre-training on large-scale heterogeneous graph and propose a novel pre-training GNN framework, named PT-HGNN. The proposed PT-HGNN designs both the node- and schema-level pre-training tasks to contrastively preserve heterogeneous semantic and structural properties as a form of transferable knowledge for various downstream tasks. In addition, a relationbased personalized PageRank is proposed to sparsify large-scale heterogeneous graph for efficient pre-training. Extensive experiments on one of the largest public heterogeneous graphs (OAG) demonstrate that our PT-HGNN significantly outperforms various state-of-the-art baselines.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper12
- Self-supervised Heterogeneous Graph Pre-training Based on Structural ClusteringYaming Yang, Ziyu Guan, Zhe Wang, Wei Zhao 等NeurIPS 2022 · 被引用 70 次
- WalkLM: A Uniform Language Model Fine-tuning Framework for Attributed Graph EmbeddingYanchao Tan, Zihao Zhou, Hang Lv, Weiming Liu 等NeurIPS 2023 · 被引用 60 次
- HetGPT: Harnessing the Power of Prompt Tuning in Pre-Trained Heterogeneous Graph Neural NetworksYihong Ma, Ning Yan, Jiayu Li, Masood S. Mortazavi 等WWW 2024 · 被引用 50 次
- Beyond Redundancy: Information-aware Unsupervised Multiplex Graph Structure LearningZhixiang Shen, Shuo Wang, Zhao KangNeurIPS 2024 · 被引用 46 次
- Link Prediction on Latent Heterogeneous GraphsTrung-Kien Nguyen, Zemin Liu, Yuan FangWWW 2023 · 被引用 14 次
它引用的顶会 Paper9
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- ALBERT: A Lite BERT for Self-supervised Learning of Language RepresentationsZhenzhong Lan, Mingda Chen, Sebastian Goodman, Kevin Gimpel 等ICLR 2020 · 被引用 7,418 次
- Graph Contrastive Learning with AugmentationsYuning You, Tianlong Chen, Yongduo Sui, Ting Chen 等NeurIPS 2020 · 被引用 3,042 次
- Understanding Contrastive Representation Learning through Alignment and Uniformity on the HypersphereTongzhou Wang, Phillip IsolaICML 2020 · 被引用 2,360 次
- Strategies for Pre-training Graph Neural NetworksWeihua Hu, Bowen Liu, Joseph Gomes, Marinka Zitnik 等ICLR 2020 · 被引用 1,744 次
相关 Paper
- GPT-GNN: Generative Pre-Training of Graph Neural NetworksZiniu Hu, Yuxiao Dong, Kuansan Wang, Kai-Wei Chang 等KDD 2020 · 被引用 438 次
- Learning to Pre-train Graph Neural NetworksYuanfu Lu, Xunqiang Jiang, Yuan Fang, Chuan ShiAAAI 2021 · 被引用 158 次
- Non-Homophilic Graph Pre-Training and Prompt LearningXingtong Yu, Jie Zhang, Yuan Fang, Renhe JiangKDD 2025 · 被引用 6 次
- Harnessing Language Model for Cross-Heterogeneity Graph Knowledge TransferJinyu Yang, Ruijia Wang, Cheng Yang, Bo Yan 等AAAI 2025 · 被引用 4 次
- MUG: Meta-path-aware Universal Heterogeneous Graph Pre-TrainingLianze Shan, Jitao Zhao, Dongxiao He, Yongqi Huang 等AAAI 2026 · 被引用 1 次
