An Efficient Subgraph-Inferring Framework for Large-Scale Heterogeneous Graphs
Wei Zhou, Hong Huang, Ruize Shi, Kehan Yin, Hai Jin
Abstract
Heterogeneous Graph Neural Networks (HGNNs) play a vital role in advancing the field of graph representation learning by addressing the complexities arising from diverse data types and interconnected relationships in real-world scenarios. However, traditional HGNNs face challenges when applied to large-scale graphs due to the necessity of training or inferring on the entire graph. As the size of the heterogeneous graphs increases, the time and memory overhead required by these models escalates rapidly, even reaching unacceptable levels. To address this issue, in this paper, we present a novel framework named (SubInfer), which conducts training and inferring on subgraphs instead of the entire graphs, hence efficiently handling large-scale heterogeneous graphs. The proposed framework comprises three main steps: 1) partitioning the heterogeneous graph from multiple perspectives to preserve various semantic information, 2) completing the subgraphs to improve the convergence speed of subgraph training and the performance of subgraph inference, and 3) training and inferring the HGNN model on distributed clusters to further reduce the time overhead. The framework is applicable to the vast majority of HGNN models. Experiments on five benchmark datasets demonstrate that SubInfer effectively optimizes the training and inference phase, delivering comparable performance to traditional HGNN models while significantly reducing time and memory overhead.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f948bb2f-d0c6-48cc-abfb-e45b2cb113a2Builds on5
- LightGCN: Simplifying and Powering Graph Convolution Network for RecommendationXiangnan He, Kuan Deng, Xiang Wang, Yan Li et al.SIGIR 2020 · 4,448 citations
- GraphSAINT: Graph Sampling Based Inductive Learning MethodHanqing Zeng, Hongkuan Zhou, Ajitesh Srivastava, Rajgopal Kannan et al.ICLR 2020 · 1,155 citations
- Are we really making much progress?: Revisiting, benchmarking and refining heterogeneous graph neural networksQingsong Lv, Ming Ding, Qiang Liu, Yuxiang Chen et al.KDD 2021 · 249 citations
- Simple and Efficient Heterogeneous Graph Neural NetworkXiaocheng Yang, Mingyu Yan, Shirui Pan, Xiaochun Ye et al.AAAI 2023 · 233 citations
- xGCN: An Extreme Graph Convolutional Network for Large-scale Social Link PredictionXiran Song, Jianxun Lian, Hong Huang, Zihan Luo et al.WWW 2023 · 33 citations
Related papers
- NeutronHeter: Optimizing Distributed Graph Neural Network Training for Heterogeneous ClustersChunyu Cao, Xin Ai, Qiange Wang, Yanfeng Zhang et al.SIGMOD 2026 · 3 citations
- Pre-training on Large-Scale Heterogeneous GraphXunqiang Jiang, Tianrui Jia, Yuan Fang, Chuan Shi et al.KDD 2021 · 44 citations
- FastGL: A GPU-Efficient Framework for Accelerating Sampling-Based GNN Training at Large ScaleZeyu Zhu, Peisong Wang, Qinghao Hu, Gang Li et al.ASPLOS 2024 · 8 citations
- A Flexible, Equivariant Framework for Subgraph GNNs via Graph Products and Graph CoarseningGuy Bar-Shalom, Yam Eitan, Fabrizio Frasca, Haggai MaronNeurIPS 2024 · 9 citations
- Hardware/Software Co-Programmable Framework for Computational SSDs to Accelerate Deep Learning Service on Large-Scale GraphsMiryeong Kwon, Donghyun Gouk, Sangwon Lee, Myoungsoo JungFAST 2022 · 32 citations
