Graph2Video: Leveraging Video Models to Model Dynamic Graph Evolution
Hua Liu, Yanbin Wei, Fei Xing, Tyler Derr, Haoyu Han, Yu Zhang
Abstract
Dynamic graphs are common in real-world systems such as social media, recommender systems, and traffic networks. Existing dynamic graph models for link prediction often fall short in capturing the complexity of temporal evolution. They tend to overlook fine-grained variations in temporal interaction order, struggle with dependencies that span long time horizons, and offer limited capability to model pair-specific relational dynamics. To address these challenges, we propose Graph2Video, a video-inspired framework that views the temporal neighborhood of a target link as a sequence of "graph frames". By stacking temporally ordered subgraph frames into a "graph video", Graph2Video leverages the inductive biases of video foundation models to capture both fine-grained local variations and long-range temporal dynamics. It generates a link-level embedding that serves as a lightweight and plug-and-play link-centric memory unit. This embedding integrates seamlessly into existing dynamic graph encoders, effectively addressing the limitations of prior approaches. Extensive experiments on benchmark datasets show that Graph2Video outperforms state-of-the-art baselines on the link prediction task in most cases. The results highlight the potential of borrowing spatio-temporal modeling techniques from computer vision as a promising and effective approach for advancing dynamic graph learning.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext dddf88fc-a32f-428b-90f7-fb2f74789578Builds on19
- ViViT: A Video Vision TransformerAnurag Arnab, Mostafa Dehghani, Georg Heigold, Chen Sun et al.ICCV 2021 · 2,947 citations
- Is Space-Time Attention All You Need for Video Understanding?Gedas Bertasius, Heng Wang, Lorenzo TorresaniICML 2021 · 2,927 citations
- Multiscale Vision TransformersHaoqi Fan, Bo Xiong, Karttikeya Mangalam, Yanghao Li et al.ICCV 2021 · 1,611 citations
- Inductive representation learning on temporal graphsDa Xu, Chuanwei Ruan, Evren Körpeoglu, Sushant Kumar et al.ICLR 2020 · 901 citations
- Traffic Flow Prediction via Spatial Temporal Graph Neural NetworkXiaoyang Wang, Yao Ma, Yiqi Wang, Wei Jin et al.WWW 2020 · 644 citations
Related papers
- Repeat-Aware Neighbor Sampling for Dynamic Graph LearningTao Zou, Yuhao Mao, Junchen Ye, Bowen DuKDD 2024 · 9 citations
- TAWRMAC: A Novel Dynamic Graph Representation Learning MethodSoheila Farokhi, Xiaojun Qi, Hamid KarimiWWW 2026
- FTM: A Frame-Level Timeline Modeling Method for Temporal Graph Representation LearningBowen Cao, Qichen Ye, Weiyuan Xu, Yuexian ZouAAAI 2023 · 1 citation
- Global-Lens Transformers: Adaptive Token Mixing for Dynamic Link PredictionTao Zou, Chengfeng Wu, Tianxi Liao, Junchen Ye et al.AAAI 2026
- SALoM: Structure Aware Temporal Graph Networks with Long-Short Memory UpdaterHanwen Liu, Longjiao Zhang, Rui Wang, Tongya Zheng et al.NeurIPS 2025 · 6 citations
