Re-thinking Knowledge Graph Completion Evaluation from an Information Retrieval Perspective
Ying Zhou, Xuanang Chen, Ben He, Zheng Ye, Le Sun
Abstract
Knowledge graph completion (KGC) aims to infer missing knowledge triples based on known facts in a knowledge graph. Current KGC research mostly follows an entity ranking protocol, wherein the effectiveness is measured by the predicted rank of a masked entity in a test triple. The overall performance is then given by a micro(-average) metric over all individual answer entities. Due to the incomplete nature of the large-scale knowledge bases, such an entity ranking setting is likely affected by unlabelled top-ranked positive examples, raising questions on whether the current evaluation protocol is sufficient to guarantee a fair comparison of KGC systems. To this end, this paper presents a systematic study on whether and how the label sparsity affects the current KGC evaluation with the popular micro metrics. Specifically, inspired by the TREC paradigm for large-scale information retrieval (IR) experimentation, we create a relatively "complete" judgment set based on a sample from the popular FB15k-237 dataset following the TREC pooling method. According to our analysis, it comes as a surprise that switching from the original labels to our "complete" labels results in a drastic change of system ranking of a variety of 13 popular KGC models in terms of micro metrics. Further investigation indicates that the IR-like macro(-average) metrics are more stable and discriminative under different settings, meanwhile, less affected by label sparsity. Thus, for KGC evaluation, we recommend conducting TREC-style pooling to balance between human efforts and label completeness, and reporting also the IR-like macro metrics to reflect the ranking nature of the KGC task.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 7d2d0650-a5e1-466e-bef3-6463975d1e68Cited by top-tier papers2
- DPCL-Diff:Temporal Knowledge Graph Reasoning Based on Graph Node Diffusion Model with Dual-Domain Periodic Contrastive LearningYukun Cao, Lisheng Wang, Luobin HuangAAAI 2025 · 5 citations
- DuetGraph: Coarse-to-Fine Knowledge Graph Reasoning with Dual-Pathway Global-Local FusionJin Li, Zezhong Ding, Xike XieNeurIPS 2025 · 5 citations
Builds on14
- Composition-based Multi-Relational Graph Convolutional NetworksShikhar Vashishth, Soumya Sanyal, Vikram Nitin, Partha P. TalukdarICLR 2020 · 1,105 citations
- Learning Hierarchy-Aware Knowledge Graph Embeddings for Link PredictionZhanqiu Zhang, Jianyu Cai, Yongdong Zhang, Jie WangAAAI 2020 · 481 citations
- InteractE: Improving Convolution-Based Knowledge Graph Embeddings by Increasing Feature InteractionsShikhar Vashishth, Soumya Sanyal, Vikram Nitin, Nilesh Agrawal et al.AAAI 2020 · 393 citations
- You CAN Teach an Old Dog New Tricks! On Training Knowledge Graph EmbeddingsDaniel Ruffinelli, Samuel Broscheit, Rainer GemullaICLR 2020 · 238 citations
- HittER: Hierarchical Transformers for Knowledge Graph EmbeddingsSanxing Chen, Xiaodong Liu, Jianfeng Gao, Jian Jiao et al.EMNLP 2021 · 110 citations
Related papers
- Rethinking Knowledge Graph Evaluation Under the Open-World AssumptionHaotong Yang, Zhouchen Lin, Muhan ZhangNeurIPS 2022 · 30 citations
- Revisiting the Evaluation Protocol of Knowledge Graph Completion Methods for Link PredictionSudhanshu Tiwari, Iti Bansal, Carlos R. RiveroWWW 2021 · 16 citations
- Are We Wasting Time? A Fast, Accurate Performance Evaluation Framework for Knowledge Graph Link PredictorsFilip Cornell, Yifei Jin, Jussi Karlgren, Sarunas GirdzijauskasICDE 2025 · 1 citation
- KRACL: Contrastive Learning with Graph Context Modeling for Sparse Knowledge Graph CompletionZhaoxuan Tan, Zilong Chen, Shangbin Feng, Qingyue Zhang et al.WWW 2023 · 50 citations
- Towards Global-Topology Relation Graph for Inductive Knowledge Graph CompletionLing Ding, Lei Huang, Zhizhi Yu, Di Jin et al.AAAI 2025 · 8 citations
