Serving Graph Compression for Graph Neural Networks
Si Si, Felix X. Yu, Ankit Singh Rawat, Cho-Jui Hsieh, Sanjiv Kumar
Abstract
Serving a GNN model online is challenging --- in many applications when testing nodes are connected to training nodes, one has to propagate information from training nodes to testing nodes to achieve the best performance, and storing the whole training set (including training graph and node features) during inference stage is prohibitive for large-scale problems. In this paper, we study graph compression to reduce the storage requirement for GNN in serving. Given a GNN model to be served, we propose to construct a compressed graph with a smaller number of nodes. In serving time, one just needs to replace the original training set graph by this compressed graph, without the need of changing the actual GNN model and the forward pass. We carefully analyze the error in the forward pass and derive simple ways to construct the compressed graph to minimize the approximation error. Experimental results on semi-supervised node classification demonstrate that the proposed method can significantly reduce the serving space requirement for GNN inference.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 9290659e-7e65-4fbf-beae-1eef46ced47bCited by top-tier papers4
- Graph Condensation for Inductive Node Representation LearningXinyi Gao, Tong Chen, Yilong Zang, Wentao Zhang et al.ICDE 2024 · 32 citations
- Rethinking and Accelerating Graph Condensation: A Training-Free Approach with Class PartitionXinyi Gao, Guanhua Ye, Tong Chen, Wentao Zhang et al.WWW 2025 · 27 citations
- TEDDY: Trimming Edges with Degree-based Discrimination StrategyHyunjin Seo, Jihun Yun, Eunho YangICLR 2024 · 3 citations
- Towards Pre-trained Graph Condensation via Optimal TransportYeyu Yan, Shuai Zheng, Wenjun Hui, Xiangkai Zhu et al.NeurIPS 2025 · 3 citations
Related papers
- TT-GNN: Efficient On-Chip Graph Neural Network Training via Embedding Reformation and Hardware OptimizationZheng Qu, Dimin Niu, Shuangchen Li, Hongzhong Zheng et al.MICRO 2023 · 6 citations
- Graph Condensation for Graph Neural NetworksWei Jin, Lingxiao Zhao, Shichang Zhang, Yozen Liu et al.ICLR 2022 · 203 citations
- SC-GNN: A Communication-Efficient Semantic Compression for Distributed Training of GNNsJihe Wang, Ying Wu, Danghui WangDAC 2024 · 2 citations
- Joint Edge-Model Sparse Learning is Provably Efficient for Graph Neural NetworksShuai Zhang, Meng Wang, Pin-Yu Chen, Sijia Liu et al.ICLR 2023
- Eliminating Data Processing Bottlenecks in GNN Training over Large Graphs via Two-level Feature CompressionYuxin Ma, Ping Gong, Tianming Wu, Jiawei Yi et al.VLDB 2024 · 10 citations
