Attention Mechanisms Perspective: Exploring LLM Processing of Graph-Structured Data
Zhong Guan, Likang Wu, Hongke Zhao, Ming He, Jianping Fan
摘要
Attention mechanisms are critical to the success of large language models (LLMs), driving significant advancements in multiple fields. However, for graph-structured data, which requires emphasis on topological connections, they fall short compared to message-passing mechanisms on fixed links, such as those employed by Graph Neural Networks (GNNs). This raises a question: "Does attention fail for graphs in natural language settings?" Motivated by these observations, we embarked on an empirical study from the perspective of attention mechanisms to explore how LLMs process graph-structured data. The goal is to gain deeper insights into the attention behavior of LLMs over graph structures. We uncovered unique phenomena regarding how LLMs apply attention to graph-structured data and analyzed these findings to improve the modeling of such data by LLMs. The primary findings of our research are: 1) While LLMs can recognize graph data and capture text-node interactions, they struggle to model inter-node relationships within graph structures due to inherent architectural constraints. 2) The attention distribution of LLMs across graph nodes does not align with ideal structural patterns, indicating a failure to adapt to graph topology nuances. 3) Neither fully connected attention nor fixed connectivity is optimal; each has specific limitations in its application scenarios. Instead, intermediate-state attention windows improve LLM training performance and seamlessly transition to fully connected windows during inference. Source code: LLM4Exploration
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Actions Speak Louder than Prompts: A Large-Scale Study of LLMs for Graph InferenceBen Finkelshtein, Silviu Cucerzan, Sujay Kumar Jauhar, Ryen W WhiteICLR 2026 · 被引用 5 次
- Formalizing and Mitigating Structural Distortion in LLM Attention for Graph ReasoningDonald Loveland, Puja Trivedi, Ari Weinstein, Edward W. Huang 等KDD 2026
- CITE: Benchmarking Heterogeneous Text-Attributed Graph ModelsChenghao Zhang, Qingqing Long, Ludi Wang, Wenjuan Cui 等ACL 2026
- D-RAG: Differentiable Retrieval-Augmented Generation for Knowledge Graph Question AnsweringGuangze Gao, Zixuan Li, Chunfeng Yuan, Jiawei Li 等EMNLP 2025
- CoT is Not the Chain of Truth: An Empirical Internal Analysis of Reasoning LLMs for Fake News GenerationZhao Tong, Chunlin Gong, Yiping Zhang, Haichao Shi 等ICML 2026
它引用的顶会 Paper15
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida 等NeurIPS 2022 · 被引用 24,707 次
- Efficient Streaming Language Models with Attention SinksGuangxuan Xiao, Yuandong Tian, Beidi Chen, Song Han 等ICLR 2024 · 被引用 1,714 次
- Pure Transformers are Powerful Graph LearnersJinwoo Kim, Dat Nguyen, Seonwoo Min, Sungjun Cho 等NeurIPS 2022 · 被引用 311 次
- GraphFormers: GNN-nested Transformers for Representation Learning on Textual GraphJunhan Yang, Zheng Liu, Shitao Xiao, Chaozhuo Li 等NeurIPS 2021 · 被引用 262 次
相关 Paper
- Talk like a Graph: Encoding Graphs for Large Language ModelsBahare Fatemi, Jonathan Halcrow, Bryan PerozziICLR 2024 · 被引用 194 次
- Efficient Code Analysis via Graph Representation Learning-Guided Large Language ModelsHang Gao, Tao Peng, Baoquan Cui, Hong Huang 等ICML 2026 · 被引用 1 次
- SLASH the Sink: Sharpening Structural Attention Inside LLMsYiming Liu, Bin Lu, Xinbing Wang, Chenghu Zhou 等ICML 2026
- Digest the Knowledge: Large Language Models empowered Message Passing for Knowledge Graph Question AnsweringJunhong Wan, Tao Yu, Kunyu Jiang, Yao Fu 等ACL 2025 · 被引用 4 次
- Can Graph Learning Improve Planning in LLM-based Agents?Xixi Wu, Yifei Shen, Caihua Shan, Kaitao Song 等NeurIPS 2024 · 被引用 67 次
