A Graph is Worth K Words: Euclideanizing Graph using Pure Transformer
Zhangyang Gao, Daize Dong, Cheng Tan, Jun Xia, Bozhen Hu, Stan Z. Li
Abstract
Can we model Non-Euclidean graphs as pure language or even Euclidean vectors while retaining their inherent information? The Non-Euclidean property have posed a long term challenge in graph modeling. Despite recent graph neural networks and graph transformers efforts encoding graphs as Euclidean vectors, recovering the original graph from vectors remains a challenge. In this paper, we introduce GraphsGPT, featuring an Graph2Seq encoder that transforms Non-Euclidean graphs into learnable Graph Words in the Euclidean space, along with a GraphGPT decoder that reconstructs the original graph from Graph Words to ensure information equivalence. We pretrain GraphsGPT on M molecules and yield some interesting findings: (1) The pretrained Graph2Seq excels in graph representation learning, achieving state-of-the-art results on graph classification and regression tasks. (2) The pretrained GraphGPT serves as a strong graph generator, demonstrated by its strong ability to perform both few-shot and conditional graph generation. (3) Graph2Seq+GraphGPT enables effective graph mixup in the Euclidean space, overcoming previously known Non-Euclidean challenges. (4) The edge-centric pretraining framework GraphsGPT demonstrates its efficacy in graph domain tasks, excelling in both representation and generation. Code is available at GitHub.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext bbe1a732-abc1-4ee6-a153-ab5185657f22Cited by top-tier papers3
- AlphaFold Database Debiasing for Robust Inverse FoldingCheng Tan, Zhenxiao Cao, Zhangyang Gao, Siyuan Li et al.NeurIPS 2025 · 3 citations
- Graph Generative Pre-trained TransformerXiaohui Chen, Yinkai Wang, Jiaxing He, Yuanqi Du et al.ICML 2025
- Adaptive Recurrent Message Passing for Test Time Computing on GraphsJunshu Sun, Wanxing Chang, Qingming Huang, Shuhui WangICML 2026
Builds on53
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Flamingo: a Visual Language Model for Few-Shot LearningJean-Baptiste Alayrac, Jeff Donahue, Pauline Luc, Antoine Miech et al.NeurIPS 2022 · 6,707 citations
- Graph Contrastive Learning with AugmentationsYuning You, Tianlong Chen, Yongduo Sui, Ting Chen et al.NeurIPS 2020 · 3,042 citations
Related papers
- GraphGPT: Generative Pre-trained Graph Eulerian TransformerQifang Zhao, Weidong Ren, Tianyu Li, Hong Liu et al.ICML 2025
- Towards A Universal Graph Structural EncoderJialin Chen, Haolan Zuo, Haoyu Wang, Siqi Miao et al.WWW 2026 · 6 citations
- Geometric Graph Representation Learning on Protein Structure PredictionTian Xia, Wei-Shinn KuKDD 2021 · 28 citations
- AutoGT: Automated Graph Transformer Architecture SearchZizhao Zhang, Xin Wang, Chaoyu Guan, Ziwei Zhang et al.ICLR 2023
- View Space: Learning Representation across Arbitrary GraphsDooho Lee, Myeong Kong, Minho Jeong, Jaemin YooICML 2026 · 2 citations
