Serving Graph Compression for Graph Neural Networks
Si Si, Felix X. Yu, Ankit Singh Rawat, Cho-Jui Hsieh, Sanjiv Kumar
摘要
Serving a GNN model online is challenging --- in many applications when testing nodes are connected to training nodes, one has to propagate information from training nodes to testing nodes to achieve the best performance, and storing the whole training set (including training graph and node features) during inference stage is prohibitive for large-scale problems. In this paper, we study graph compression to reduce the storage requirement for GNN in serving. Given a GNN model to be served, we propose to construct a compressed graph with a smaller number of nodes. In serving time, one just needs to replace the original training set graph by this compressed graph, without the need of changing the actual GNN model and the forward pass. We carefully analyze the error in the forward pass and derive simple ways to construct the compressed graph to minimize the approximation error. Experimental results on semi-supervised node classification demonstrate that the proposed method can significantly reduce the serving space requirement for GNN inference.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper4
- Graph Condensation for Inductive Node Representation LearningXinyi Gao, Tong Chen, Yilong Zang, Wentao Zhang 等ICDE 2024 · 被引用 32 次
- Rethinking and Accelerating Graph Condensation: A Training-Free Approach with Class PartitionXinyi Gao, Guanhua Ye, Tong Chen, Wentao Zhang 等WWW 2025 · 被引用 27 次
- TEDDY: Trimming Edges with Degree-based Discrimination StrategyHyunjin Seo, Jihun Yun, Eunho YangICLR 2024 · 被引用 3 次
- Towards Pre-trained Graph Condensation via Optimal TransportYeyu Yan, Shuai Zheng, Wenjun Hui, Xiangkai Zhu 等NeurIPS 2025 · 被引用 3 次
相关 Paper
- TT-GNN: Efficient On-Chip Graph Neural Network Training via Embedding Reformation and Hardware OptimizationZheng Qu, Dimin Niu, Shuangchen Li, Hongzhong Zheng 等MICRO 2023 · 被引用 6 次
- Graph Condensation for Graph Neural NetworksWei Jin, Lingxiao Zhao, Shichang Zhang, Yozen Liu 等ICLR 2022 · 被引用 203 次
- SC-GNN: A Communication-Efficient Semantic Compression for Distributed Training of GNNsJihe Wang, Ying Wu, Danghui WangDAC 2024 · 被引用 2 次
- Joint Edge-Model Sparse Learning is Provably Efficient for Graph Neural NetworksShuai Zhang, Meng Wang, Pin-Yu Chen, Sijia Liu 等ICLR 2023
- Eliminating Data Processing Bottlenecks in GNN Training over Large Graphs via Two-level Feature CompressionYuxin Ma, Ping Gong, Tianming Wu, Jiawei Yi 等VLDB 2024 · 被引用 10 次
