LMC: Fast Training of GNNs via Subgraph Sampling with Provable Convergence
Zhihao Shi, Xize Liang, Jie Wang
Abstract
The message passing-based graph neural networks (GNNs) have achieved great success in many real-world applications. However, training GNNs on large-scale graphs suffers from the well-known neighbor explosion problem, i.e., the exponentially increasing dependencies of nodes with the number of message passing layers. Subgraph-wise sampling methods -- a promising class of mini-batch training techniques -- discard messages outside the mini-batches in backward passes to avoid the neighbor explosion problem at the expense of gradient estimation accuracy. This poses significant challenges to their convergence analysis and convergence speeds, which seriously limits their reliable real-world applications. To address this challenge, we propose a novel subgraph-wise sampling method with a convergence guarantee, namely Local Message Compensation (LMC). To the best of our knowledge, LMC is the first subgraph-wise sampling method with provable convergence. The key idea of LMC is to retrieve the discarded messages in backward passes based on a message passing formulation of backward passes. By efficient and effective compensations for the discarded messages in both forward and backward passes, LMC computes accurate mini-batch gradients and thus accelerates convergence. We further show that LMC converges to first-order stationary points of GNNs. Experiments on large-scale benchmark tasks demonstrate that LMC significantly outperforms state-of-the-art subgraph-wise sampling methods in terms of efficiency.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 50a700ea-0388-4983-900e-efda67731a9bCited by top-tier papers20
- VQGraph: Rethinking Graph Representation Space for Bridging GNNs and MLPsLing Yang, Ye Tian, Minkai Xu, Zhongyi Liu et al.ICLR 2024 · 48 citations
- Layer-Neighbor Sampling - Defusing Neighborhood Explosion in GNNsMuhammed Fatih Balin, Ümit V. ÇatalyürekNeurIPS 2023 · 37 citations
- AFDGCF: Adaptive Feature De-correlation Graph Collaborative Filtering for RecommendationsWei Wu, Chao Wang, Dazhong Shen, Chuan Qin et al.SIGIR 2024 · 33 citations
- A Circuit Domain Generalization Framework for Efficient Logic Synthesis in Chip DesignZhihai Wang, Lei Chen, Jie Wang, Yinqi Bai et al.ICML 2024 · 13 citations
- Mixture of Weak and Strong Experts on GraphsHanqing Zeng, Hanjia Lyu, Diyi Hu, Yinglong Xia et al.ICLR 2024 · 11 citations
Builds on10
- Open Graph Benchmark: Datasets for Machine Learning on GraphsWeihua Hu, Matthias Fey, Marinka Zitnik, Yuxiao Dong et al.NeurIPS 2020 · 3,935 citations
- Simple and Deep Graph Convolutional NetworksMing Chen, Zhewei Wei, Zengfeng Huang, Bolin Ding et al.ICML 2020 · 1,910 citations
- GraphSAINT: Graph Sampling Based Inductive Learning MethodHanqing Zeng, Hongkuan Zhou, Ajitesh Srivastava, Rajgopal Kannan et al.ICLR 2020 · 1,155 citations
- Decoupling the Depth and Scope of Graph Neural NetworksHanqing Zeng, Muhan Zhang, Yinglong Xia, Ajitesh Srivastava et al.NeurIPS 2021 · 189 citations
- Implicit Graph Neural NetworksFangda Gu, Heng Chang, Wenwu Zhu, Somayeh Sojoudi et al.NeurIPS 2020 · 188 citations
Related papers
- Faster Convergence in Mini-batch Graph Neural Networks Training with Pseudo Full Neighborhood CompensationQiqi Zhou, Yanyan Shen, Lei ChenVLDB 2025 · 2 citations
- Accurate and Scalable Graph Neural Networks via Message InvarianceZhihao Shi, Jie Wang, Zhiwei Zhuang, Xize Liang et al.ICLR 2025
- VQ-GNN: A Universal Framework to Scale up Graph Neural Networks using Vector QuantizationMucong Ding, Kezhi Kong, Jingling Li, Chen Zhu et al.NeurIPS 2021 · 68 citations
- GCN meets GPU: Decoupling "When to Sample" from "How to Sample"Morteza Ramezani, Weilin Cong, Mehrdad Mahdavi, Anand Sivasubramaniam et al.NeurIPS 2020 · 37 citations
- Resource-Efficient Training for Large Graph Convolutional Networks with Label-Centric Cumulative SamplingMingkai Lin, Wenzhong Li, Ding Li, Yizhou Chen et al.WWW 2022 · 10 citations
