Supercharging Graph Transformers with Advective Diffusion
Qitian Wu, Chenxiao Yang, Kaipeng Zeng, Michael M. Bronstein
Abstract
The capability of generalization is a cornerstone for the success of modern learning systems. For non-Euclidean data, e.g., graphs, that particularly involves topological structures, one important aspect neglected by prior studies is how machine learning models generalize under topological shifts. This paper proposes ADVDIFFORMER, a physics-inspired graph Transformer model designed to address this challenge. The model is derived from advective diffusion equations which describe a class of continuous message passing process with observed and latent topological structures. We show that ADVDIFFORMER has provable capability for controlling generalization error with topological shifts, which in contrast cannot be guaranteed by graph diffusion models, i.e., the generalization of common graph neural networks in continuous space. Empirically, the model demonstrates superiority in various predictive tasks across information networks, molecular screening and protein interactions 1 .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3c931469-40f4-4e70-8a1e-ce69fa1749e7Cited by top-tier papers2
- Decoupling Universal Laws and Environmental Heterogeneity: A Physics-Inspired Framework for Robust Spatio-Temporal ForecastingAoyu Liu, Liming Wei, YAYING ZHANGICML 2026
- Geometric Graph Neural Diffusion for Stable Molecular Dynamics SimulationsHaokai Hong, Wanyu Lin, Zhang Chusong, KC TanICLR 2026
Builds on23
- Open Graph Benchmark: Datasets for Machine Learning on GraphsWeihua Hu, Matthias Fey, Marinka Zitnik, Yuxiao Dong et al.NeurIPS 2020 · 3,935 citations
- WILDS: A Benchmark of in-the-Wild Distribution ShiftsPang Wei Koh, Shiori Sagawa, Henrik Marklund, Sang Michael Xie et al.ICML 2021 · 1,773 citations
- Recipe for a General, Powerful, Scalable Graph TransformerLadislav Rampásek, Michael Galkin, Vijay Prakash Dwivedi, Anh Tuan Luu et al.NeurIPS 2022 · 1,216 citations
- Understanding over-squashing and bottlenecks on graphs via curvatureJake Topping, Francesco Di Giovanni, Benjamin Paul Chamberlain, Xiaowen Dong et al.ICLR 2022 · 628 citations
- NodeFormer: A Scalable Graph Structure Learning Transformer for Node ClassificationQitian Wu, Wentao Zhao, Zenan Li, David P. Wipf et al.NeurIPS 2022 · 472 citations
Related papers
- DiGress: Discrete Denoising diffusion for graph generationClément Vignac, Igor Krawczuk, Antoine Siraudin, Bohan Wang et al.ICLR 2023 · 70 citations
- Learning Divergence Fields for Shift-Robust Graph RepresentationsQitian Wu, Fan Nie, Chenxiao Yang, Junchi YanICML 2024 · 6 citations
- Equiformer: Equivariant Graph Attention Transformer for 3D Atomistic GraphsYi-Lun Liao, Tess E. SmidtICLR 2023 · 65 citations
- DIFFormer: Scalable (Graph) Transformers Induced by Energy Constrained DiffusionQitian Wu, Chenxiao Yang, Wentao Zhao, Yixuan He et al.ICLR 2023 · 26 citations
- Invariant Graph Transformer for Out-of-Distribution GeneralizationTianyin Liao, Ziwei Zhang, Yufei Sun, Chunyu Hu et al.KDD 2026 · 2 citations
