WalkLM: A Uniform Language Model Fine-tuning Framework for Attributed Graph Embedding
Yanchao Tan, Zihao Zhou, Hang Lv, Weiming Liu, Carl Yang
Abstract
Graphs are widely used to model interconnected entities and improve downstream predictions in various real-world applications. However, real-world graphs nowadays are often associated with complex attributes on multiple types of nodes and even links that are hard to model uniformly, while the widely used graph neural networks (GNNs) often require sufficient training toward specific downstream predictions to achieve strong performance. In this work, we take a fundamentally different approach than GNNs, to simultaneously achieve deep joint modeling of complex attributes and flexible structures of real-world graphs and obtain unsupervised generic graph representations that are not limited to specific downstream predictions. Our framework, built on a natural integration of language models (LMs) and random walks (RWs), is straightforward, powerful and data-efficient. Specifically, we first perform attributed RWs on the graph and design an automated program to compose roughly meaningful textual sequences directly from the attributed RWs; then we fine-tune an LM using the RW-based textual sequences and extract embedding vectors from the LM, which encapsulates both attribute semantics and graph structures. In our experiments, we evaluate the learned node embeddings towards different downstream prediction tasks on multiple real-world attributed graph datasets and observe significant improvements over a comprehensive set of state-of-the-art unsupervised node embedding methods. We believe this work opens a door for more sophisticated technical designs and empirical evaluations toward the leverage of LMs for the modeling of real-world graphs.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b15d591b-2322-41e0-8359-a7bf4d2bc218Cited by top-tier papers16
- G-Refer: Graph Retrieval-Augmented Large Language Model for Explainable RecommendationYuhan Li, Xinni Zhang, Linhao Luo, Heng Chang et al.WWW 2025 · 46 citations
- RAGraph: A General Retrieval-Augmented Graph Learning FrameworkXinke Jiang, Rihong Qiu, Yongxin Xu, Wentao Zhang et al.NeurIPS 2024 · 42 citations
- Large Language Model Meets Graph Neural Network in Knowledge DistillationShengxiang Hu, Guobing Zou, Song Yang, Shiyi Lin et al.AAAI 2025 · 19 citations
- ZeroG: Investigating Cross-dataset Zero-shot Transferability in GraphsYuhan Li, Peisong Wang, Zhixun Li, Jeffrey Xu Yu et al.KDD 2024 · 19 citations
- GRAVER: Generative Graph Vocabularies for Robust Graph Foundation Models Fine-tuningHaonan Yuan, Qingyun Sun, Junhua Shi, Xingcheng Fu et al.NeurIPS 2025 · 17 citations
Builds on24
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida et al.NeurIPS 2022 · 24,707 citations
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu et al.ICLR 2022 · 18,833 citations
- Graph Contrastive Learning with Adaptive AugmentationYanqiao Zhu, Yichen Xu, Feng Yu, Qiang Liu et al.WWW 2021 · 1,415 citations
- MAGNN: Metapath Aggregated Graph Neural Network for Heterogeneous Graph EmbeddingXinyu Fu, Jiani Zhang, Ziqiao Meng, Irwin KingWWW 2020 · 1,149 citations
Related papers
- LGA: LLM-GNN Aggregation for Temporal Evolution Attribute Graph PredictionFeng Zhao, Ruoyu Chai, Kangzheng Liu, Xianggan LiuEMNLP 2025
- Graph Language ModelsMoritz Plenz, Anette FrankACL 2024
- Leveraging Large Language Models for Node Generation in Few-Shot Learning on Text-Attributed GraphsJianxiang Yu, Yuxiang Ren, Chenghua Gong, Jiaqi Tan et al.AAAI 2025 · 32 citations
- GPT-GNN: Generative Pre-Training of Graph Neural NetworksZiniu Hu, Yuxiao Dong, Kuansan Wang, Kai-Wei Chang et al.KDD 2020 · 438 citations
- Graph-R1: Incentivizing the Zero-Shot Graph Learning Capability in LLMs via Explicit ReasoningYicong Wu, Guangyue Lu, Yuan Zuo, Huarong Zhang et al.EMNLP 2025
