Can LLMs be Good Graph Judge for Knowledge Graph Construction?
Haoyu Huang, Chong Chen, Zeang Sheng, Yang Li, Wentao Zhang
Abstract
In real-world scenarios, most of the data obtained from the information retrieval (IR) system is unstructured. Converting natural language sentences into structured Knowledge Graphs (KGs) remains a critical challenge. We identified three limitations with respect to existing KG construction methods: (1) There could be a large amount of noise in real-world documents, which could result in extracting messy information. (2) Naive LLMs usually extract inaccurate knowledge from some domainspecific documents. (3) Hallucination phenomenon cannot be overlooked when directly using LLMs to construct KGs. In this paper, we propose GraphJudge, a KG construction framework to address the aforementioned challenges. In this framework, we designed an entity-centric strategy to eliminate the noise information in the documents. And we fine-tuned a LLM as a graph judge to finally enhance the quality of generated KGs. Experiments conducted on two general and one domain-specific text-graph pair datasets demonstrate state-ofthe-art performance against various baseline methods with strong generalization abilities. Our code is available at https://github.com/hhyhuang/GraphJudge .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ad129549-dd1c-40e1-ba72-56ce405f16d7Cited by top-tier papers4
- AtlasKV: Augmenting LLMs with Billion-Scale Knowledge Graphs in 20GB VRAMHaoyu Huang, Hong Ting Tsang, Jiaxin Bai, Xi Peng et al.ICLR 2026 · 4 citations
- Template-Theorems Graph Construction to Enhance Mathematical Reasoning Capabilities of LLMYarong Lan, Yajing Xu, Huajun ChenAAAI 2026
- Hyper-KGGen: A Skill-Driven Knowledge Extractor for High-Quality Knowledge Hypergraph GenerationRizhuo Huang, Yifan Feng, Rundong Xue, Shihui Ying et al.KDD 2026
- LLMs as Knowledge Graph Refiners: Mitigating Factual Inconsistencies in Generative Knowledge ExtractionDonghyun Kim, Hyeongjun Yang, Seokju Hwang, Kyong-Ho Lee et al.ACL 2026
Builds on10
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu et al.ICLR 2022 · 18,833 citations
- Large Language Models Can Be Easily Distracted by Irrelevant ContextFreda Shi, Xinyun Chen, Kanishka Misra, Nathan Scales et al.ICML 2023 · 970 citations
- Large language models are few-shot clinical information extractorsMonica Agrawal, Stefan Hegselmann, Hunter Lang, Yoon Kim et al.EMNLP 2022 · 285 citations
- GPT-RE: In-context Learning for Relation Extraction using Large Language ModelsZhen Wan, Fei Cheng, Zhuoyuan Mao, Qianying Liu et al.EMNLP 2023 · 132 citations
- Revisiting DocRED - Addressing the False Negative Problem in Relation ExtractionQingyu Tan, Lu Xu, Lidong Bing, Hwee Tou Ng et al.EMNLP 2022 · 76 citations
Related papers
- KnowGPT: Knowledge Graph based Prompting for Large Language ModelsQinggang Zhang, Junnan Dong, Hao Chen, Daochen Zha et al.NeurIPS 2024 · 66 citations
- Discovering Latent Facts from Context to Construct Richer Open Knowledge GraphsJinpeng Li, Hang Yu, Ziqi Ma, Peng QiAAAI 2026
- Improving the Robustness of Knowledge-Grounded Dialogue via Contrastive LearningJiaan Wang, Jianfeng Qu, Kexin Wang, Zhixu Li et al.AAAI 2024 · 5 citations
- Scaling Knowledge Graph Construction through Synthetic Data Generation and DistillationPrafulla Kumar Choubey, Xin Su, Man Luo, XIANGYU PENG et al.ICLR 2026 · 5 citations
- MKGL: Mastery of a Three-Word LanguageLingbing Guo, Zhongpu Bo, Zhuo Chen, Yichi Zhang et al.NeurIPS 2024 · 27 citations
