Typing Errors in Factual Knowledge Graphs: Severity and Possible Ways Out
Peiran Yao, Denilson Barbosa
摘要
Large-scale factual knowledge graphs (KGs) such as DBpedia and Wikidata are essential to many popular downstream tasks and are also widely used by various research communities as training and/or benchmarking data. Despite their immense success and utility, these KGs are surprisingly noisy. In this study, we investigate the quality of these KGs, where the typing error rate is estimated to be 27% for coarse-grained types on average, and even 73% for certain fine-grained types. In pursuit of solutions, we propose an active typing error detection algorithm that maximizes the utilization of both gold and noisy labels. We also comprehensively discuss and compare the state-of-the-art in unsupervised, semi-supervised, and supervised paradigms to deal with typing errors in factual KGs. The outcomes of this study provide guidelines for researchers to use noisy factual KGs. To help practitioners deploy the techniques and conduct further research, we published our code and data 1.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper1
相关 Paper
- An End-To-End Re-Evaluation of Table Entity-LinkersMartin Pekár Christensen, Matteo Lissandrini, Katja HoseICDE 2026
- Knowledge Graph Error Detection with Contrastive Confidence AdaptionXiangyu Liu, Yang Liu, Wei HuAAAI 2024 · 被引用 16 次
- Extraction of Validating Shapes from very large Knowledge GraphsKashif Rabbani, Matteo Lissandrini, Katja HoseVLDB 2023 · 被引用 48 次
- REA: Robust Cross-lingual Entity Alignment Between Knowledge GraphsShichao Pei, Lu Yu, Guoxian Yu, Xiangliang ZhangKDD 2020 · 被引用 44 次
- PGE: Robust Product Graph Embedding Learning for Error DetectionKewei Cheng, Xian Li, Yifan Ethan Xu, Xin Luna Dong 等VLDB 2022 · 被引用 13 次
