SUREL+: Moving from Walks to Sets for Scalable Subgraph-based Graph Representation Learning
Haoteng Yin, Muhan Zhang, Jianguo Wang, Pan Li
Abstract
Subgraph-based graph representation learning (SGRL) has recently emerged as a powerful tool in many prediction tasks on graphs due to its advantages in model expressiveness and generalization ability. Most previous SGRL models face computational issues related to the high cost of extracting subgraphs for each training or testing query. Recently, SUREL was proposed to accelerate SGRL, which samples random walks offline and joins these walks online as a proxy of subgraphs for prediction. Thanks to the reusability of sampled walks across different queries, SUREL achieves state-of-the-art performance in terms of scalability and prediction accuracy. However, SUREL still suffers from high computational overhead caused by node redundancy in sampled walks. In this work, we propose a novel framework SUREL+ that upgrades SUREL by using node sets instead of walks to represent subgraphs. By definition, such set-based representations avoid repeated nodes, but node sets can be irregular in size. To solve this issue, we design a dedicated sparse data structure to efficiently store and access node sets, and provide a specialized operator to join them in parallel batches. SUREL+ is modularized to support multiple types of set samplers, structural features, and neural encoders to complement the loss of structural information after the reduction from walks to sets. Extensive experiments have been performed to verify the effectiveness of SUREL+ in the prediction tasks of links, relation types, and higher-order patterns. SUREL+ achieves 3--11× speedups of SUREL while maintaining comparable or even better prediction performance; compared to other SGRL baselines, SUREL+ achieves 20× speedups and significantly improves the prediction accuracy.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 7e4f90b2-15aa-482f-a379-edd5ba996bccCited by top-tier papers10
- Revisiting Link Prediction: a data perspectiveHaitao Mao, Juanhui Li, Harry Shomer, Bingheng Li et al.ICLR 2024 · 40 citations
- Structural Information Enhanced Graph Representation for Link PredictionLei Shi, Bin Hu, Deng Zhao, Jianshan He et al.AAAI 2024 · 21 citations
- Cross-links Matter for Link Prediction: Rethinking the Debiased GNN from a Data PerspectiveZihan Luo, Hong Huang, Jianxun Lian, Xiran Song et al.NeurIPS 2023 · 17 citations
- Breaking the Entanglement of Homophily and Heterophily in Semi-supervised Node ClassificationHenan Sun, Xunkai Li, Zhengyu Wu, Daohan Su et al.ICDE 2024 · 9 citations
- GENTI: GPU-powered Walk-based Subgraph Extraction for Scalable Representation Learning on Dynamic GraphsZihao Yu, Ningyi Liao, Siqiang LuoVLDB 2024 · 8 citations
Builds on23
- Open Graph Benchmark: Datasets for Machine Learning on GraphsWeihua Hu, Matthias Fey, Marinka Zitnik, Yuxiao Dong et al.NeurIPS 2020 · 3,935 citations
- GraphSAINT: Graph Sampling Based Inductive Learning MethodHanqing Zeng, Hongkuan Zhou, Ajitesh Srivastava, Rajgopal Kannan et al.ICLR 2020 · 1,155 citations
- Inductive Relation Prediction by Subgraph ReasoningKomal K. Teru, Etienne G. Denis, William L. HamiltonICML 2020 · 493 citations
- Can Graph Neural Networks Count Substructures?Zhengdao Chen, Lei Chen, Soledad Villar, Joan BrunaNeurIPS 2020 · 392 citations
- Distance Encoding: Design Provably More Powerful Neural Networks for Graph Representation LearningPan Li, Yanbang Wang, Hongwei Wang, Jure LeskovecNeurIPS 2020 · 391 citations
Related papers
- Algorithm and System Co-design for Efficient Subgraph-based Graph Representation LearningHaoteng Yin, Muhan Zhang, Yanbang Wang, Jianguo Wang et al.VLDB 2022 · 47 citations
- SG-Serve: Efficient Model Serving for Subgraph-based Graph Representation LearningQihui Zhou, Peiqi Yin, Xiao Yan, Changji Li et al.SIGMOD 2026
- TGL: A General Framework for Temporal GNN Training onBillion-Scale GraphsHongkuan Zhou, Da Zheng, Israt Nisa, Vassilis N. Ioannidis et al.VLDB 2022 · 109 citations
- Translating Subgraphs to Nodes Makes Simple GNNs Strong and Efficient for Subgraph Representation LearningDongkwan Kim, Alice OhICML 2024 · 6 citations
- Multi-Level Graph Representation Learning Through Predictive Community-based PartitioningBo-Young Lim, Jeongha Park, Kisung Lee, Hyuk-Yoon KwonSIGMOD 2025 · 2 citations
