Realistic Re-evaluation of Knowledge Graph Completion Methods: An Experimental Study
Farahnaz Akrami, Mohammed Samiul Saeef, Qingheng Zhang, Wei Hu, Chengkai Li
Abstract
In the active research area of employing embedding models for knowledge graph completion, particularly for the task of link prediction, most prior studies used two benchmark datasets FB15k and WN18 in evaluating such models. Most triples in these and other datasets in such studies belong to reverse and duplicate relations which exhibit high data redundancy due to semantic duplication, correlation or data incompleteness. This is a case of excessive data leakage---a model is trained using features that otherwise would not be available when the model needs to be applied for real prediction. There are also Cartesian product relations for which every triple formed by the Cartesian product of applicable subjects and objects is a true fact. Link prediction on the aforementioned relations is easy and can be achieved with even better accuracy using straightforward rules instead of sophisticated embedding models. A more fundamental defect of these models is that the link prediction scenario, given such data, is non-existent in the real-world. This paper is the first systematic study with the main objective of assessing the true effectiveness of embedding models when the unrealistic triples are removed. Our experiment results show these models are much less accurate than what we used to perceive. Their poor accuracy renders link prediction a task without truly effective automated solution. Hence, we call for re-investigation of possible effective approaches.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0c5568f3-20c9-4d6f-a76a-e05c4db71754Cited by top-tier papers23
- A Benchmarking Study of Embedding-based Entity Alignment for Knowledge GraphsZequn Sun, Qingheng Zhang, Wei Hu, Chengming Wang et al.VLDB 2020 · 297 citations
- CoDEx: A Comprehensive Knowledge Graph Completion BenchmarkTara Safavi, Danai KoutraEMNLP 2020 · 97 citations
- MulDE: Multi-teacher Knowledge Distillation for Low-dimensional Knowledge Graph EmbeddingsKai Wang, Yu Liu, Qian Ma, Quan Z. ShengWWW 2021 · 67 citations
- Explaining Link Prediction Systems based on Knowledge Graph EmbeddingsAndrea Rossi, Donatella Firmani, Paolo Merialdo, Tommaso TeofiliSIGMOD 2022 · 47 citations
- Efficient Non-Sampling Knowledge Graph EmbeddingZelong Li, Jianchao Ji, Zuohui Fu, Yingqiang Ge et al.WWW 2021 · 41 citations
Related papers
- Revisiting the Evaluation Protocol of Knowledge Graph Completion Methods for Link PredictionSudhanshu Tiwari, Iti Bansal, Carlos R. RiveroWWW 2021 · 16 citations
- A Semantic Filter Based on Relations for Knowledge Graph CompletionZongwei Liang, Junan Yang, Hui Liu, Ke-Ju HuangEMNLP 2021 · 6 citations
- Can We Predict New Facts with Open Knowledge Graph Embeddings? A Benchmark for Open Link PredictionSamuel Broscheit, Kiril Gashteovski, Yanjie Wang, Rainer GemullaACL 2020 · 27 citations
- Are Embedded Potatoes Still Vegetables? On the Limitations of WordNet Embeddings for Lexical SemanticsXuyou Cheng, Michael Sejr Schlichtkrull, Guy EmersonEMNLP 2023
- ExpressivE: A Spatio-Functional Embedding For Knowledge Graph CompletionAleksandar Pavlovic, Emanuel SallingerICLR 2023 · 12 citations
