Accelerating Scalable Graph Neural Network Inference with Node-Adaptive Propagation
Xinyi Gao, Wentao Zhang, Junliang Yu, Yingxia Shao, Quoc Viet Hung Nguyen, Bin Cui, Hongzhi Yin
Abstract
Graph neural networks (GNNs) have exhibited exceptional efficacy in a diverse array of applications. However, the sheer size of large-scale graphs presents a significant challenge to real-time inference with GNNs. Although existing Scalable GNNs leverage linear propagation to preprocess the features and accelerate the training and inference procedure, these methods still suffer from scalability issues when making inferences on unseen nodes, as the feature preprocessing requires the graph to be known and fixed. To further accelerate Scalable GNNs inference in this inductive setting, we propose an online propagation framework and two novel node-adaptive propagation methods that can customize the optimal propagation depth for each node based on its topological information and thereby avoid redundant feature propagation. The trade-off between accuracy and latency can be flexibly managed through simple hyper-parameters to accommodate various latency constraints. Moreover, to compensate for the inference accuracy loss caused by the potential early termination of propagation, we further propose Inception Distillation to exploit the multi-scale receptive field information within graphs. The rigorous and comprehensive experimental study on public datasets with varying scales and characteristics demonstrates that the proposed inference acceleration framework outperforms existing state-of-the-art graph inference acceleration methods in terms of accuracy and efficiency. Particularly, the superiority of our approach is notable on datasets with larger scales, yielding ainference speedup on the largest Ogbn-products dataset.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4275a107-c445-4f3f-b6eb-ae716293d0f1Cited by top-tier papers6
- Rethinking and Accelerating Graph Condensation: A Training-Free Approach with Class PartitionXinyi Gao, Guanhua Ye, Tong Chen, Wentao Zhang et al.WWW 2025 · 27 citations
- Apt-Serve: Adaptive Request Scheduling on Hybrid Cache for Scalable LLM Inference ServingShihong Gao, Xin Zhang, Yanyan Shen, Lei ChenSIGMOD 2025 · 7 citations
- A Comprehensive Benchmark on Spectral GNNs: The Impact on Efficiency, Memory, and EffectivenessNingyi Liao, Haoyu Liu, Zulun Zhu, Siqiang Luo et al.SIGMOD 2026 · 4 citations
- Relational Database Distillation: From Structured Tables to Condensed Graph DataXinyi Gao, Jingxi Zhang, Lijian Chen, Tong Chen et al.WWW 2026 · 2 citations
- Contrastive Graph Condensation: Advancing Data Versatility through Self-Supervised LearningXinyi Gao, Yayong Li, Tong Chen, Guanhua Ye et al.KDD 2025 · 1 citation
Builds on25
- Open Graph Benchmark: Datasets for Machine Learning on GraphsWeihua Hu, Matthias Fey, Marinka Zitnik, Yuxiao Dong et al.NeurIPS 2020 · 3,935 citations
- GraphSAINT: Graph Sampling Based Inductive Learning MethodHanqing Zeng, Hongkuan Zhou, Ajitesh Srivastava, Rajgopal Kannan et al.ICLR 2020 · 1,155 citations
- Simple Spectral Graph ConvolutionHao Zhu, Piotr KoniuszICLR 2021 · 352 citations
- Graph-less Neural Networks: Teaching Old MLPs New Tricks Via DistillationShichang Zhang, Yozen Liu, Yizhou Sun, Neil ShahICLR 2022 · 234 citations
- A Unified Lottery Ticket Hypothesis for Graph Neural NetworksTianlong Chen, Yongduo Sui, Xuxi Chen, Aston Zhang et al.ICML 2021 · 208 citations
Related papers
- Rethinking Node-wise Propagation for Large-scale Graph LearningXunkai Li, Jingyuan Ma, Zhengyu Wu, Daohan Su et al.WWW 2024 · 21 citations
- Graph Condensation for Inductive Node Representation LearningXinyi Gao, Tong Chen, Yilong Zang, Wentao Zhang et al.ICDE 2024 · 32 citations
- Accelerating Large Scale Real-Time GNN Inference using Channel PruningHongkuan Zhou, Ajitesh Srivastava, Hanqing Zeng, Rajgopal Kannan et al.VLDB 2021 · 86 citations
- NodeBits: A Plug-and-Play Framework for Accelerating Graph Inference by Post-Hoc Binary QuantizationQihao Cheng, Tianhao Wu, Da Yan, Haoran TangKDD 2026
- SCARA: Scalable Graph Neural Networks with Feature-Oriented OptimizationNingyi Liao, Dingheng Mo, Siqiang Luo, Xiang Li et al.VLDB 2022 · 36 citations
