LLMs as Stratification Signals for KG Accuracy Evaluation
Stefano Marchesin, Matteo Ceccarello, Gianmaria Silvello
摘要
Knowledge Graph (KG) accuracy assessment is essential for ensuring data quality in downstream applications, yet remains prohibitively expensive due to annotation costs and scale. Large Language Models (LLMs), trained on vast corpora, offer cheap fact validation but remain unreliable as direct accuracy estimators due to hallucinations and knowledge gaps. We propose a novel approach that exploits LLM capabilities without relying on their correctness: using aggregated LLM predictions as stratification signals for sampling-based accuracy estimation. By partitioning KGs into internally homogeneous strata guided by aggregated LLM outputs, we achieve statistically significant cost reductions ranging from 11% to 54% over unstratified and topology-based baselines on real-world KGs. To scale beyond LLM computational constraints, we introduce a knowledge distillation strategy that transfers stratification signals to efficient student models, requiring annotation of only 0.25% of facts while maintaining signal quality. Experiments on six KGs spanning 20M+ triples demonstrate consistent improvements over SotA methods, with statistical guarantees on accuracy estimates.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper2
相关 Paper
- HOLMES: Hyper-Relational Knowledge Graphs for Multi-hop Question Answering using LLMsPranoy Panda, Ankush Agarwal, Chaitanya Devaguptapu, Manohar Kaul 等ACL 2024
- Large Language Model Meets Graph Neural Network in Knowledge DistillationShengxiang Hu, Guobing Zou, Song Yang, Shiyi Lin 等AAAI 2025 · 被引用 19 次
- Evidence-Focused Fact Summarization for Knowledge-Augmented Zero-Shot Question AnsweringSungho Ko, Hyunjin Cho, Hyungjoo Chae, Jinyoung Yeo 等EMNLP 2024 · 被引用 4 次
- Digest the Knowledge: Large Language Models empowered Message Passing for Knowledge Graph Question AnsweringJunhong Wan, Tao Yu, Kunyu Jiang, Yao Fu 等ACL 2025 · 被引用 4 次
- Knowledge Reasoning Language Model: Unifying Knowledge and Language for Inductive Knowledge Graph ReasoningXingrui Zhuo, Jiapu Wang, Gongqing Wu, Zhongyuan Wang 等ICLR 2026 · 被引用 2 次
