AutoGT: Automated Graph Transformer Architecture Search
Zizhao Zhang, Xin Wang, Chaoyu Guan, Ziwei Zhang, Haoyang Li, Wenwu Zhu
摘要
Although Transformer architectures have been successfully applied to graph data with the advent of Graph Transformer, current design of Graph Transformer still heavily relies on human labor and expertise knowledge to decide proper neural architectures and suitable graph encoding strategies at each Transformer layer. In literature, there have been some works on automated design of Transformers focusing on non-graph data such as texts and images without considering graph encoding strategies, which fail to handle the non-euclidean graph data. In this paper, we study the problem of automated graph Transformer, for the first time. However, solving these problems poses the following challenges: i) how can we design a unified search space for graph Transformer, and ii) how to deal with the coupling relations between Transformer architectures and the graph encodings of each Transformer layer. To address these challenges, we propose Automated Graph Transformer (AutoGT), a neural architecture search framework that can automatically discover the optimal graph Transformer architectures by joint optimization of Transformer architecture and graph encoding strategies. Specifically, we first propose a unified graph Transformer formulation that can represent most of state-of-the-art graph Transformer architectures. Based upon the unified formulation, we further design the graph Transformer search space that includes both candidate architectures and various graph encodings. To handle the coupling relations, we propose a novel encoding-aware performance estimation strategy by gradually training and splitting the supernets according to the correlations between graph encodings and architectures. The proposed strategy can provide a more consistent and fine-grained performance prediction when evaluating the jointly optimized graph encodings and architectures. Extensive experiments and ablation studies show that our proposed AutoGT gains sufficient improvement over state-of-the-art hand-crafted baselines on all datasets, demonstrating its effectiveness and wide applicability.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper9
- Adaptive Disentangled Transformer for Sequential RecommendationYipeng Zhang, Xin Wang, Hong Chen, Wenwu ZhuKDD 2023 · 被引用 32 次
- Multi-task Graph Neural Architecture Search with Task-aware Collaboration and CurriculumYijian Qin, Xin Wang, Ziwei Zhang, Hong Chen 等NeurIPS 2023 · 被引用 27 次
- What Improves the Generalization of Graph Transformers? A Theoretical Dive into the Self-attention and Positional EncodingHongkang Li, Meng Wang, Tengfei Ma, Sijia Liu 等ICML 2024 · 被引用 23 次
- Customized Subgraph Selection and Encoding for Drug-drug Interaction PredictionHaotong Du, Quanming Yao, Juzheng Zhang, Yang Liu 等NeurIPS 2024 · 被引用 23 次
- TorchGT: A Holistic System for Large-Scale Graph Transformer TrainingMeng Zhang, Jie Sun, Qinghao Hu, Peng Sun 等SC 2024 · 被引用 7 次
相关 Paper
- AutoGEL: An Automated Graph Neural Network with Explicit Link InformationZhili Wang, Shimin Di, Lei ChenNeurIPS 2021 · 被引用 46 次
- Dynamic Heterogeneous Graph Attention Neural Architecture SearchZeyang Zhang, Ziwei Zhang, Xin Wang, Yijian Qin 等AAAI 2023 · 被引用 44 次
- AutoGSR: Neural Architecture Search for Graph-based Session RecommendationJingfan Chen, Guanghui Zhu, Haojun Hou, Chunfeng Yuan 等SIGIR 2022 · 被引用 25 次
- Graph External Attention Enhanced TransformerJianqing Liang, Min Chen, Jiye LiangICML 2024 · 被引用 11 次
- Hypergraph Neural Architecture SearchWei Lin, Xu Peng, Zhengtao Yu, Taisong JinAAAI 2024 · 被引用 3 次
