PRES: Toward Scalable Memory-Based Dynamic Graph Neural Networks
Junwei Su, Difan Zou, Chuan Wu
摘要
Memory-based Dynamic Graph Neural Networks (MDGNNs) are a family of dynamic graph neural networks that leverage a memory module to extract, distill, and memorize long-term temporal dependencies, leading to superior performance compared to memory-less counterparts. However, training MDGNNs faces the challenge of handling entangled temporal and structural dependencies, requiring sequential and chronological processing of data sequences to capture accurate temporal patterns. During the batch training, the temporal data points within the same batch will be processed in parallel, while their temporal dependencies are neglected. This issue is referred to as temporal discontinuity and restricts the effective temporal batch size, limiting data parallelism and reducing MDGNNs' flexibility in industrial applications. This paper studies the efficient training of MDGNNs at scale, focusing on the temporal discontinuity in training MDGNNs with large temporal batch sizes. We first conduct a theoretical study on the impact of temporal batch size on the convergence of MDGNN training. Based on the analysis, we propose PRES, an iterative prediction-correction scheme combined with a memory coherence learning objective to mitigate the effect of temporal discontinuity, enabling MDGNNs to be trained with significantly larger temporal batches without sacrificing generalization performance. Experimental results demonstrate that our approach enables up to a 4 × larger temporal batch (3.4× speed-up) during MDGNN training.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- Towards Robust Graph Incremental Learning on Evolving GraphsJunwei Su, Difan Zou, Zijun Zhang, Chuan WuICML 2023 · 被引用 37 次
- MSPipe: Efficient Temporal GNN Training via Staleness-Aware PipelineGuangming Sheng, Junwei Su, Chao Huang, Chuan WuKDD 2024 · 被引用 7 次
- Temporal-Aware Evaluation and Learning for Temporal Graph Neural NetworksJunwei Su, Shan WuAAAI 2025 · 被引用 3 次
- Unlocking Multi-Modal Potentials for Link Prediction on Dynamic Text-Attributed GraphsYuanyuan Xu, Wenjie Zhang, Ying Zhang, Xuemin Lin 等AAAI 2026 · 被引用 2 次
- TIDFormer: Exploiting Temporal and Interactive Dynamics Makes A Great Dynamic Graph TransformerJie Peng, Zhewei Wei, Yuhang YeKDD 2025 · 被引用 2 次
它引用的顶会 Paper14
- EvolveGCN: Evolving Graph Convolutional Networks for Dynamic GraphsAldo Pareja, Giacomo Domeniconi, Jie Chen, Tengfei Ma 等AAAI 2020 · 被引用 1,429 次
- Large Batch Optimization for Deep Learning: Training BERT in 76 minutesYang You, Jing Li, Sashank J. Reddi, Jonathan Hseu 等ICLR 2020 · 被引用 1,170 次
- Inductive representation learning on temporal graphsDa Xu, Chuanwei Ruan, Evren Körpeoglu, Sushant Kumar 等ICLR 2020 · 被引用 901 次
- Don't Use Large Mini-batches, Use Local SGDTao Lin, Sebastian U. Stich, Kumar Kshitij Patel, Martin JaggiICLR 2020 · 被引用 462 次
- Is Local SGD Better than Minibatch SGD?Blake E. Woodworth, Kumar Kshitij Patel, Sebastian U. Stich, Zhen Dai 等ICML 2020 · 被引用 277 次
相关 Paper
- DistTGL: Distributed Memory-Based Temporal Graph Neural Network TrainingHongkuan Zhou, Da Zheng, Xiang Song, George Karypis 等SC 2023 · 被引用 21 次
- Effective and Efficient Distributed Temporal Graph Learning through Hotspot Memory SharingLongjiao Zhang, Rui Wang, Tongya Zheng, Ziqi Huang 等VLDB 2025 · 被引用 1 次
- SEIGN: A Simple and Efficient Graph Neural Network for Large Dynamic GraphsXiao Qin, Nasrullah Sheikh, Chuan Lei, Berthold Reinwald 等ICDE 2023 · 被引用 14 次
- PipeTGL: (Near) Zero Bubble Memory-based Temporal Graph Neural Network Training via Pipeline OptimizationJun Liu, Bingqian Du, Ziyue Luo, Sitian Lu 等VLDB 2025
- PiPAD: Pipelined and Parallel Dynamic GNN Training on GPUsChunyang Wang, Desen Sun, Yuebin BaiPPoPP 2023 · 被引用 27 次
