HiGPT: Heterogeneous Graph Language Model
Jiabin Tang, Yuhao Yang, Wei Wei, Lei Shi, Long Xia, Dawei Yin, Chao Huang
Abstract
Heterogeneous graph learning aims to capture complex relationships and diverse relational semantics among entities in a heterogeneous graph to obtain meaningful representations for nodes and edges. Recent advancements in heterogeneous graph neural networks (HGNNs) have achieved state-of-the-art performance by considering relation heterogeneity and using specialized message functions and aggregation rules. However, existing frameworks for heterogeneous graph learning have limitations in generalizing across diverse heterogeneous graph datasets. Most of these frameworks follow the "pre-train" and "fine-tune" paradigm on the same dataset, which restricts their capacity to adapt to new and unseen data. This raises the question: "Can we generalize heterogeneous graph models to be well-adapted to diverse downstream learning tasks with distribution shifts in both node token sets and relation type heterogeneity?" To tackle those challenges, we propose HiGPT, a general large graph model with <u>H</u>eterogeneous graph <u>i</u>nstruction-tuning paradigm. Our framework enables learning from arbitrary heterogeneous graphs without the need for any fine-tuning process from downstream datasets. To handle distribution shifts in heterogeneity, we introduce an in-context heterogeneous graph tokenizer that captures semantic relationships in different heterogeneous graphs, facilitating model adaptation. We incorporate a large corpus of heterogeneity-aware graph instructions into our HiGPT, enabling the model to effectively comprehend complex relation heterogeneity and distinguish between various types of graph tokens. Furthermore, we introduce the Mixture-of-Thought (MoT) instruction augmentation paradigm to mitigate data scarcity by generating diverse and informative instructions. Through comprehensive evaluations conducted in various settings, our proposed framework demonstrates exceptional performance in terms of generalization performance, surpassing current leading benchmarks. We make our model implementation openly available, along with comprehensive details at: https://github.com/HKUDS/HiGPT.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0f6e7520-ff55-4f2c-aa60-6fb844058581Cited by top-tier papers25
- SelfGNN: Self-Supervised Graph Neural Networks for Sequential RecommendationYuxi Liu, Lianghao Xia, Chao HuangSIGIR 2024 · 62 citations
- GFM-RAG: Graph Foundation Model for Retrieval Augmented GenerationLinhao Luo, Zicheng Zhao, Reza Haffari, Dinh Phung et al.NeurIPS 2025 · 54 citations
- PromptMM: Multi-Modal Knowledge Distillation for Recommendation with Prompt-TuningWei Wei, Jiabin Tang, Lianghao Xia, Yangqin Jiang et al.WWW 2024 · 46 citations
- SAMGPT: Text-free Graph Foundation Model for Multi-domain Pre-training and Cross-domain AdaptationXingtong Yu, Zechuan Gong, Chang Zhou, Yuan Fang et al.WWW 2025 · 45 citations
- GRAVER: Generative Graph Vocabularies for Robust Graph Foundation Models Fine-tuningHaonan Yuan, Qingyun Sun, Junhua Shi, Xingcheng Fu et al.NeurIPS 2025 · 17 citations
Builds on21
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma et al.NeurIPS 2022 · 22,562 citations
- Tree of Thoughts: Deliberate Problem Solving with Large Language ModelsShunyu Yao, Dian Yu, Jeffrey Zhao, Izhak Shafran et al.NeurIPS 2023 · 5,068 citations
- MAGNN: Metapath Aggregated Graph Neural Network for Heterogeneous Graph EmbeddingXinyu Fu, Jiani Zhang, Ziqiao Meng, Irwin KingWWW 2020 · 1,149 citations
- Rethinking the Role of Demonstrations: What Makes In-Context Learning Work?Sewon Min, Xinxi Lyu, Ari Holtzman, Mikel Artetxe et al.EMNLP 2022 · 634 citations
Related papers
- HetGPT: Harnessing the Power of Prompt Tuning in Pre-Trained Heterogeneous Graph Neural NetworksYihong Ma, Ning Yan, Jiayu Li, Masood S. Mortazavi et al.WWW 2024 · 50 citations
- Instruction-based Hypergraph PretrainingMingdai Yang, Zhiwei Liu, Liangwei Yang, Xiaolong Liu et al.SIGIR 2024 · 4 citations
- HePa: Heterogeneous Graph Prompting for All-Level Classification TasksJia Jinghong, Lei Song, Jiaxing Li, Youyong KongAAAI 2025 · 2 citations
- Pre-training on Large-Scale Heterogeneous GraphXunqiang Jiang, Tianrui Jia, Yuan Fang, Chuan Shi et al.KDD 2021 · 44 citations
- GraphGPT: Graph Instruction Tuning for Large Language ModelsJiabin Tang, Yuhao Yang, Wei Wei, Lei Shi et al.SIGIR 2024 · 182 citations
