AutoTransfer: AutoML with Knowledge Transfer - An Application to Graph Neural Networks
Kaidi Cao, Jiaxuan You, Jiaju Liu, Jure Leskovec
Abstract
AutoML has demonstrated remarkable success in finding an effective neural architecture for a given machine learning task defined by a specific dataset and an evaluation metric. However, most present AutoML techniques consider each task independently from scratch, which requires exploring many architectures, leading to high computational costs. Here we propose AUTOTRANSFER, an AutoML solution that improves search efficiency by transferring the prior architectural design knowledge to the novel task of interest. Our key innovation includes a task-model bank that captures the model performance over a diverse set of GNN architectures and tasks, and a computationally efficient task embedding that can accurately measure the similarity among different tasks. Based on the task-model bank and the task embeddings, we estimate the design priors of desirable models of the novel task, by aggregating a similarity-weighted sum of the top-K design distributions on tasks that are similar to the task of interest. The computed design priors can be used with any AutoML search algorithm. We evaluate AUTOTRANSFER on six datasets in the graph machine learning domain. Experiments demonstrate that (i) our proposed task embedding can be computed efficiently, and that tasks with similar embeddings have similar best-performing architectures; (ii) AUTOTRANSFER significantly improves search efficiency with the transferred design priors, reducing the number of explored architectures by an order of magnitude. Finally, we release GNN-BANK-101, a large-scale dataset of detailed GNN training information of 120,000 task-model combinations to facilitate and inspire future research.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2e1cc531-bc63-46ef-99b5-e35612402bffCited by top-tier papers4
- ExPT: Synthetic Pretraining for Few-Shot Experimental DesignTung Nguyen, Sudhanshu Agrawal, Aditya GroverNeurIPS 2023 · 27 citations
- A Pre-training Framework for Relational Data with Information-theoretic PrinciplesQuang Truong, Zhikai Chen, Mingxuan Ju, Tong Zhao et al.NeurIPS 2025 · 4 citations
- Relatron: Automating Relational Machine Learning over Relational DatabasesZhikai Chen, Han Xie, Jian Zhang, Jiliang Tang et al.ICLR 2026 · 2 citations
- Beyond Model Base Retrieval: Weaving Knowledge to Master Fine-grained Neural Network DesignJialiang Wang, Hanmo Liu, Shimin Di, Zhili Wang et al.ICML 2026
Builds on14
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- LightGCN: Simplifying and Powering Graph Convolution Network for RecommendationXiangnan He, Kuan Deng, Xiang Wang, Yan Li et al.SIGIR 2020 · 4,448 citations
- Open Graph Benchmark: Datasets for Machine Learning on GraphsWeihua Hu, Matthias Fey, Marinka Zitnik, Yuxiao Dong et al.NeurIPS 2020 · 3,935 citations
- Learning to Simulate Complex Physics with Graph NetworksAlvaro Sanchez-Gonzalez, Jonathan Godwin, Tobias Pfaff, Rex Ying et al.ICML 2020 · 1,439 citations
- Learning Mesh-Based Simulation with Graph NetworksTobias Pfaff, Meire Fortunato, Alvaro Sanchez-Gonzalez, Peter W. BattagliaICLR 2021 · 1,175 citations
Related papers
- Design Space for Graph Neural NetworksJiaxuan You, Zhitao Ying, Jure LeskovecNeurIPS 2020 · 409 citations
- Structuring Benchmark into Knowledge Graphs to Assist Large Language Models in Retrieving and Designing ModelsHanmo Liu, Shimin Di, Jialiang Wang, Zhili Wang et al.ICLR 2025
- Arch-Graph: Acyclic Architecture Relation Predictor for Task-Transferable Neural Architecture SearchMinbin Huang, Zhijian Huang, Changlin Li, Xin Chen et al.CVPR 2022 · 20 citations
- AutoGEL: An Automated Graph Neural Network with Explicit Link InformationZhili Wang, Shimin Di, Lei ChenNeurIPS 2021 · 46 citations
- FastGNAS: Accelerating and Scaling Graph Neural Architecture Search on Multi-GPUs via Ring-Based Model MigrationZhen Song, Hao Li, Tianyi Li, Yu Gu et al.SIGMOD 2026
