Resource-Efficient Training for Large Graph Convolutional Networks with Label-Centric Cumulative Sampling
Mingkai Lin, Wenzhong Li, Ding Li, Yizhou Chen, Sanglu Lu
摘要
Graph Convolutional Networks (GCNs) are popular for learning representation of graph data and have a wide range of applications in social networks, recommendation systems, etc. However, training GCN models for large networks is resource intensive and time consuming, which hinders them from real deployment. The existing GCN training methods intended to optimize the sampling of mini-batches for stochastic gradient descent to accelerate training process, which did not reduce the problem size and had limited reduction in computation complexity. In this paper, we argue that a GCN can be trained with a sampled subgraph to produce approximate node representations, which inspires us a novel perspective to accelerate GCN training via network sampling. To this end, we propose a label-centric cumulative sampling (LCS) framework for training GCNs for large graphs. The proposed method constructs a subgraph cumulatively based on probabilistic sampling, and trains the GCN model iteratively to generate approximate node representations. The optimality of LCS is theoretically guaranteed to minimize the bias during node aggregation procedure in GCN training. Extensive experiments based on four real-world network datasets show that the LCS framework accelerates the training for the state-of-the-art GCN models up to 17x without causing noteworthy model accuracy drop.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper6
- GraphGPT: Graph Instruction Tuning for Large Language ModelsJiabin Tang, Yuhao Yang, Wei Wei, Lei Shi 等SIGIR 2024 · 被引用 182 次
- Unified Graph Neural Networks Pre-training for Multi-domain GraphsMingkai Lin, Xiaobin Hong, Wenzhong Li, Sanglu LuAAAI 2025 · 被引用 4 次
- Contextual Structure Knowledge Transfer for Graph Neural NetworksZhiyuan Yu, Wenzhong Li, Zhangyue Yin, Xiaobin Hong 等AAAI 2025 · 被引用 3 次
- Learnable Matrix Profile for Motif Discovery on Multivariate Time SeriesMingkai Lin, Yinke Wang, Xiaobin Hong, Wenzhong LiAAAI 2026
- Demystifying GNN-to-MLP Knowledge Transfer: Theoretical Grounding and Dual-Stream Distillation MethodZhiyuan Yu, Mingkai Lin, Wenzhong Li, Zhangyue Yin 等AAAI 2026
相关 Paper
- GCN meets GPU: Decoupling "When to Sample" from "How to Sample"Morteza Ramezani, Weilin Cong, Mehrdad Mahdavi, Anand Sivasubramaniam 等NeurIPS 2020 · 被引用 37 次
- GraphSAINT: Graph Sampling Based Inductive Learning MethodHanqing Zeng, Hongkuan Zhou, Ajitesh Srivastava, Rajgopal Kannan 等ICLR 2020 · 被引用 1,155 次
- L2-GCN: Layer-Wise and Learned Efficient Training of Graph Convolutional NetworksYuning You, Tianlong Chen, Zhangyang Wang, Yang ShenCVPR 2020
- LMC: Fast Training of GNNs via Subgraph Sampling with Provable ConvergenceZhihao Shi, Xize Liang, Jie WangICLR 2023 · 被引用 5 次
- Global Neighbor Sampling for Mixed CPU-GPU Training on Giant GraphsJialin Dong, Da Zheng, Lin F. Yang, George KarypisKDD 2021 · 被引用 22 次
