A Fair and In-Depth Evaluation of Existing End-to-End Entity Linking Systems
Hannah Bast, Matthias Hertel, Natalie Prange
Abstract
Existing evaluations of entity linking systems often say little about how the system is going to perform for a particular application. There are two fundamental reasons for this. One is that many evaluations only use aggregate measures (like precision, recall, and F1 score), without a detailed error analysis or a closer look at the results. The other is that all of the widely used benchmarks have strong biases and artifacts, in particular: a strong focus on named entities, an unclear or missing specification of what else counts as an entity mention, poor handling of ambiguities, and an over-or underrepresentation of certain kinds of entities. We provide a more meaningful and fair in-depth evaluation of a variety of existing end-to-end entity linkers. We characterize their strengths and weaknesses and also report on reproducibility aspects. The detailed results of our evaluation can be inspected under https://elevant.cs.uni-freiburg.de/emnlp2023 . Our evaluation is based on several widely used benchmarks, which exhibit the problems mentioned above to various degrees, as well as on two new benchmarks, which address the problems mentioned above. The new benchmarks can be found under https://github.com/ad-freiburg/fair-entitylinking-benchmarks . * Author contributions are stated in Section 8. M.H. is funded by the Helmholtz Association's Initiative and Networking Fund through Helmholtz AI.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 85230a12-4fba-42c2-a03e-90894c3c0858Cited by top-tier papers1
Ask how each one uses itRelated papers
- DeepType 2: Superhuman Entity Linking, All You Need Is Type InteractionsJonathan RaimanAAAI 2022 · 9 citations
- Robustness Evaluation of Entity Disambiguation Using Prior Probes: the Case of Entity OvershadowingVera Provatorova, Samarth Bhargav, Svitlana Vakulenko, Evangelos KanoulasEMNLP 2021 · 7 citations
- Evaluating Entity Disambiguation and the Role of Popularity in Retrieval-Based NLPAnthony Chen, Pallavi Gudipati, Shayne Longpre, Xiao Ling et al.ACL 2021
- EDIN: An End-to-end Benchmark and Pipeline for Unknown Entity Discovery and IndexingNora Kassner, Fabio Petroni, Mikhail Plekhanov, Sebastian Riedel et al.EMNLP 2022 · 6 citations
- A Comprehensive Evaluation of Biomedical Entity Linking ModelsDavid Kartchner, Jennifer Deng, Shubham Lohiya, Tejasri Kopparthi et al.EMNLP 2023 · 6 citations
