Generalization Principles for Inference over Text-Attributed Graphs with Large Language Models
Haoyu Peter Wang, Shikun Liu, Rongzhe Wei, Pan Li
Abstract
Large language models (LLMs) have recently been introduced to graph learning, aiming to extend their zero-shot generalization success to tasks where labeled graph data is scarce. Among these applications, inference over text-attributed graphs (TAGs) presents unique challenges: existing methods struggle with LLMs' limited context length for processing large node neighborhoods and the misalignment between node embeddings and the LLM token space. To address these issues, we establish two key principles for ensuring generalization and derive the framework LLM-BP accordingly: (1) Unifying the attribute space with task-adaptive embeddings, where we leverage LLM-based encoders and task-aware prompting to enhance generalization of the text attribute embeddings; (2) Developing a generalizable graph information aggregation mechanism, for which we adopt belief propagation with LLM-estimated parameters that adapt across graphs. Evaluations on 11 real-world TAG benchmarks demonstrate that LLM-BP significantly outperforms existing approaches, achieving 8.10% improvement with task-conditional embeddings and an additional 1.71% gain from adaptive aggregation. The code 2 and task-adaptive embeddings 3 are publicly available.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers5
- View Space: Learning Representation across Arbitrary GraphsDooho Lee, Myeong Kong, Minho Jeong, Jaemin YooICML 2026 · 2 citations
- Weaving Graph over Tokens: Contextualizing Structured Sequences for LLMsJiaxuan Chen, Zixing Zhang, Ruijun Mao, Wei Sun et al.ICML 2026
- Toward Graph-Tokenizing Large Language Models with Reconstructive Graph Instruction TuningZhongjian Zhang, Xiao Wang, Mengmei Zhang, Jiarui Tan et al.WWW 2026
- Bridging Structure and Semantics: Uncertainty-Modulated Dual-Path Diffusion for Robust Text-Attributed Graph LearningZhizhi Yu, Jiachen Liu, Qingyu Li, Dongxiao He et al.ICML 2026
- When Do Graph Foundation Models Transfer? A Data-Centric TheoryJiajun Zhu, Ying Chen, Peihao Wang, Yixuan He et al.ICML 2026
Builds on24
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- Open Graph Benchmark: Datasets for Machine Learning on GraphsWeihua Hu, Matthias Fey, Marinka Zitnik, Yuxiao Dong et al.NeurIPS 2020 · 3,935 citations
- Beyond Homophily in Graph Neural Networks: Current Limitations and Effective DesignsJiong Zhu, Yujun Yan, Lingxiao Zhao, Mark Heimann et al.NeurIPS 2020 · 1,490 citations
- GraphMAE: Self-Supervised Masked Graph AutoencodersZhenyu Hou, Xiao Liu, Yukuo Cen, Yuxiao Dong et al.KDD 2022 · 533 citations
Related papers
- Can GNN be Good Adapter for LLMs?Xuanwen Huang, Kaiqiao Han, Yang Yang, Dezheng Bao et al.WWW 2024 · 107 citations
- Quantizing Text-attributed Graphs for Semantic-Structural IntegrationJianyuan Bo, Hao Wu, Yuan FangKDD 2025
- THGB: A Comprehensive Benchmark for Text-attributed Heterogeneous GraphsLixin Zhou, Zemin Liu, Yuan Fang, Dan Niu et al.AAAI 2026
- Compressing LLM Knowledge into Graph Representations for Text-attributed Graphs LearningRunhuai Chen, Dian Shen, Dandan Zhang, Kaihong Huang et al.ACL 2026
- UTAG: Leveraging LLM as a Unified Embedding Generator for Text-Attributed GraphsMingqian Ding, Jianjun Li, Zhiyuan Ma, Liwei Zhang et al.WWW 2026
