SAC-KG: Exploiting Large Language Models as Skilled Automatic Constructors for Domain Knowledge Graph
Hanzhu Chen, Xu Shen, Qitan Lv, Jie Wang, Xiaoqi Ni, Jieping Ye
摘要
Knowledge graphs (KGs) play a pivotal role in knowledge-intensive tasks across specialized domains, where the acquisition of precise and dependable knowledge is crucial. However, existing KG construction methods heavily rely on human intervention to attain qualified KGs, which severely hinders the practical applicability in real-world scenarios. To address this challenge, we propose a general KG construction framework, named SAC-KG, to exploit large language models (LLMs) as Skilled Automatic Constructors for domain Knowledge Graph. SAC-KG effectively involves LLMs as domain experts to generate specialized and precise multi-level KGs. Specifically, SAC-KG consists of three components: Generator, Verifier, and Pruner. For a given entity, Generator produces its relations and tails from raw domain corpora, to construct a specialized single-level KG. Verifier and Pruner then work together to ensure precision by correcting generation errors and determining whether newly produced tails require further iteration for the next-level KG. Experiments demonstrate that SAC-KG automatically constructs a domain KG at the scale of over one million nodes and achieves a precision of 89.32%, leading to a superior performance with over 20% increase in precision rate compared to existing state-of-the-art methods for the KG construction task. * Corresponding author. This work was done when Hanzhu Chen was an intern at Alibaba Cloud. GPT-KG all correct Correct (a) All correct triples extracted from the full version of SAC-KG. GPT-KG Correct Error (b) Triples generated by the full version of SAC-KG. GPT-KG w/o prompt Correct Error (c) Triples generated by SAC-KG w/o prompt . GPT-KG w/o text Correct Error (d) Triples generated by SAC-KG w/o text . GPT-KG w/o verifier Correct Error (e) Triples generated by SAC-KG w/o verif ier . (f) Triples generated by SAC-KG w/o pruner . Correct triples Wrong triples commonly called rice black-streaked dwarf disease c modes of transmission c distribution area (a) Rice disease case study for SAC-KG.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper14
- AutoSchemaKG: Autonomous Knowledge Graph Construction through Dynamic Schema Induction from Web-Scale CorporaJiaxin Bai, Wei Fan, Qi Hu, Qing Zong 等ACL 2026 · 被引用 27 次
- ReMindRAG: Low-Cost LLM-Guided Knowledge Graph Traversal for Efficient RAGYikuan Hu, Jifeng Zhu, Lanrui Tang, Chen HuangNeurIPS 2025 · 被引用 10 次
- LEGO-GraphRAG: Modularizing Graph-based Retrieval-Augmented Generation for Design Space ExplorationYukun Cao, Zengyi Gao, Zhiyang Li, Xike Xie 等VLDB 2025 · 被引用 5 次
- FinKario: Event-Enhanced Automated Construction of Financial Knowledge GraphXiang Li, Penglei Sun, Wanyun Zhou, Zikai Wei 等ACL 2026 · 被引用 4 次
- Tree-KG: An Expandable Knowledge Graph Construction Framework for Knowledge-intensive DomainsSongjie Niu, Kaisen Yang, Rui Zhao, Yichao Liu 等ACL 2025 · 被引用 4 次
它引用的顶会 Paper11
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu 等ICLR 2022 · 被引用 18,833 次
- Large Language Models are Zero-Shot ReasonersTakeshi Kojima, Shixiang Shane Gu, Machel Reid, Yutaka Matsuo 等NeurIPS 2022 · 被引用 8,168 次
- Large Language Models Can Be Easily Distracted by Irrelevant ContextFreda Shi, Xinyun Chen, Kanishka Misra, Nathan Scales 等ICML 2023 · 被引用 970 次
相关 Paper
- LLMs as Knowledge Graph Refiners: Mitigating Factual Inconsistencies in Generative Knowledge ExtractionDonghyun Kim, Hyeongjun Yang, Seokju Hwang, Kyong-Ho Lee 等ACL 2026
- Discovering Latent Facts from Context to Construct Richer Open Knowledge GraphsJinpeng Li, Hang Yu, Ziqi Ma, Peng QiAAAI 2026
- Extract, Define, Canonicalize: An LLM-based Framework for Knowledge Graph ConstructionBowen Zhang, Harold SohEMNLP 2024 · 被引用 65 次
- SciMKG: A Multimodal Knowledge Graph for Science Education with Text, Image, Video and AudioTong Lu, Zhichun Wang, Yaoyu Zhou, Yiming Guan 等AAAI 2026
- Scaling Knowledge Graph Construction through Synthetic Data Generation and DistillationPrafulla Kumar Choubey, Xin Su, Man Luo, XIANGYU PENG 等ICLR 2026 · 被引用 5 次
