Graph Neural Network Training Systems: A Performance Comparison of Full-Graph and Mini-Batch
Saurabh Bajaj, Hui Guan, Marco Serafini, Juelin Liu, Hojae Son
Abstract
Graph Neural Networks (GNNs) have gained significant attention in recent years due to their ability to learn representations of graph-structured data. Two common methods for training GNNs are mini-batch training and full-graph training. Since these two methods require different training pipelines and systems optimizations, two separate classes of GNN training systems emerged, each tailored for one method. Works that introduce systems belonging to a particular category predominantly compare them with other systems within the same category, offering limited or no comparison with systems from the other category. Some prior work also justifies its focus on one specific training method by arguing that it achieves higher accuracy than the alternative. The literature, however, has incomplete and contradictory evidence in this regard. In this paper, we provide a comprehensive empirical comparison of representative full-graph and mini-batch GNN training systems. We find that the mini-batch training systems consistently converge faster than the full-graph training ones across multiple datasets, GNN models, and system configurations. We also find that minibatch training techniques converge to similar to or often higher accuracy values than full-graph training ones, showing that minibatch sampling is not necessarily detrimental to accuracy. Our work highlights the importance of comparing systems across different classes, using time-to-accuracy rather than epoch time for performance comparison, and selecting appropriate hyperparameters for each training method separately.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4c97e2cd-faf5-4ec1-9c36-7cfaaec512eeCited by top-tier papers4
- Full-Graph vs. Mini-Batch Training: Comprehensive Analysis from a Batch Size and Fan-Out Size PerspectiveMengfan Liu, Da Zheng, Junwei Su, Chuan WuICLR 2026 · 2 citations
- Bingo: Radix-based Bias Factorization for Random Walk on Dynamic GraphsPinhuan Wang, Chengying Huan, Zhibin Wang, Chen Tian et al.EuroSys 2025 · 2 citations
- SNI-GNN: SmartNIC-Assisted Full-Graph GNN Training with In-Network Embedding PredictionGuofan Yu, Sitian Chen, Zhenheng Tang, Xiaowen Chu et al.ICDE 2026
- Reducing the GPU Memory Bottleneck with Lossless Compression for MLAditya K. Kamath, Arvind Krishnamurthy, Marco Canini, Simon PeterEuroSys 2026
Builds on27
- Open Graph Benchmark: Datasets for Machine Learning on GraphsWeihua Hu, Matthias Fey, Marinka Zitnik, Yuxiao Dong et al.NeurIPS 2020 · 3,935 citations
- Simplifying and Empowering Transformers for Large-Graph RepresentationsQitian Wu, Wentao Zhao, Chenxiao Yang, Hengrui Zhang et al.NeurIPS 2023 · 318 citations
- Dorylus: Affordable, Scalable, and Accurate GNN Training with Distributed CPU Servers and Serverless ThreadsJohn Thorpe, Yifan Qiao, Jonathan Eyolfson, Shen Teng et al.OSDI 2021 · 175 citations
- DistGNN: scalable distributed training for large-scale graph neural networksMd. Vasimuddin, Sanchit Misra, Guixiang Ma, Ramanarayan Mohanty et al.SC 2021 · 110 citations
- ByteGNN: Efficient Graph Neural Network Training at Large ScaleChenguang Zheng, Hongzhi Chen, Yuxuan Cheng, Zhezheng Song et al.VLDB 2022 · 107 citations
Related papers
- ParGNN: A Scalable Graph Neural Network Training Framework on multi-GPUsJunyu Gu, Shunde Li, Rongqiang Cao, Jue Wang et al.DAC 2025
- Sylvie: 3D-Adaptive and Universal System for Large-Scale Graph Neural Network TrainingMeng Zhang, Qinghao Hu, Cheng Wan, Haozhao Wang et al.ICDE 2024 · 5 citations
- Layer-Neighbor Sampling - Defusing Neighborhood Explosion in GNNsMuhammed Fatih Balin, Ümit V. ÇatalyürekNeurIPS 2023 · 37 citations
- ADGNN: Towards Scalable GNN Training with Aggregation-Difference Aware SamplingZhen Song, Yu Gu, Tianyi Li, Qing Sun et al.SIGMOD 2024 · 9 citations
- Comprehensive Evaluation of GNN Training Systems: A Data Management PerspectiveHao Yuan, Yajiong Liu, Yanfeng Zhang, Xin Ai et al.VLDB 2024 · 8 citations
