Self-supervised Quantized Representation for Seamlessly Integrating Knowledge Graphs with Large Language Models
Qika Lin, Tianzhe Zhao, Kai He, Zhen Peng, Fangzhi Xu, Ling Huang, Jingying Ma, Mengling Feng
Abstract
Due to the presence of the natural gap between Knowledge Graph (KG) structures and the natural language, the effective integration of holistic structural information of KGs with Large Language Models (LLMs) has emerged as a significant question. To this end, we propose a two-stage framework to learn and apply quantized codes for each entity, aiming for the seamless integration of KGs with LLMs. Firstly, a self-supervised quantized representation (SSQR) method is proposed to compress both KG structural and semantic knowledge into discrete codes (i.e., tokens) that align the format of language sentences. We further design KG instruction-following data by viewing these learned codes as features to directly input to LLMs, thereby achieving seamless integration. The experiment results demonstrate that SSQR outperforms existing unsupervised quantized methods, producing more distinguishable codes. Further, the fine-tuned LLaMA2 and LLaMA3.1 also have superior performance on KG link prediction and triple classification tasks, utilizing only 16 tokens per entity instead of thousands in conventional prompting methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 32002bc9-ca7b-4634-98cb-c35a235f8dd7Cited by top-tier papers6
- CodeBrain: Bridging Decoupled Tokenizer and Multi-Scale Architecture for EEG Foundation ModelJingying Ma, Feng Wu, Qika Lin, Yucheng Xing et al.ICLR 2026 · 25 citations
- MARS: Multi-Agent Adaptive Reasoning with Socratic Guidance for Automated Prompt OptimizationJian Zhang, Zhangqi Wang, Haiping Zhu, Kangda Cheng et al.AAAI 2026 · 9 citations
- : One LLM Token for Explicit Graph Structural UnderstandingJingyao Wu, Bin Lu, Zijun Di, Xiaoying Gan et al.ICLR 2026 · 2 citations
- ReaLM: Residual Quantization Bridges Knowledge Graph Embeddings and Large Language ModelsWenbin Guo, Xin Wang, Jiaoyan Chen, Lingbing Guo et al.WWW 2026
- GS-Quant: Granular Semantic and Generative Structural Quantization for Knowledge Graph CompletionQizhuo Xie, Yunhui Liu, Yu Xing, Qianzi Hou et al.ACL 2026
Builds on20
- Composition-based Multi-Relational Graph Convolutional NetworksShikhar Vashishth, Soumya Sanyal, Vikram Nitin, Partha P. TalukdarICLR 2020 · 1,105 citations
- Think-on-Graph: Deep and Responsible Reasoning of Large Language Model on Knowledge GraphJiashuo Sun, Chengjin Xu, Lumingyuan Tang, Saizhuo Wang et al.ICLR 2024 · 247 citations
- NodePiece: Compositional and Parameter-Efficient Representations of Large Knowledge GraphsMikhail Galkin, Etienne G. Denis, Jiapeng Wu, William L. HamiltonICLR 2022 · 114 citations
- Making Large Language Models Perform Better in Knowledge Graph CompletionYichi Zhang, Zhuo Chen, Lingbing Guo, Yajing Xu et al.ACM MM 2024 · 86 citations
- Paths-over-Graph: Knowledge Graph Empowered Large Language Model ReasoningXingyu Tan, Xiaoyang Wang, Qing Liu, Xiwei Xu et al.WWW 2025 · 86 citations
Related papers
- Quantizing Text-attributed Graphs for Semantic-Structural IntegrationJianyuan Bo, Hao Wu, Yuan FangKDD 2025
- K-ON: Stacking Knowledge on the Head Layer of Large Language ModelLingbing Guo, Yichi Zhang, Zhongpu Bo, Zhuo Chen et al.AAAI 2025 · 4 citations
- LightPROF: A Lightweight Reasoning Framework for Large Language Model on Knowledge GraphTu Ao, Yanhua Yu, Yuling Wang, Yang Deng et al.AAAI 2025 · 28 citations
- MKGL: Mastery of a Three-Word LanguageLingbing Guo, Zhongpu Bo, Zhuo Chen, Yichi Zhang et al.NeurIPS 2024 · 27 citations
- A New Pipeline for Knowledge Graph Reasoning Enhanced by Large Language Models Without Fine-TuningZhongwu Chen, Long Bai, Zixuan Li, Zhen Huang et al.EMNLP 2024 · 3 citations
