Robustness Evaluation of Entity Disambiguation Using Prior Probes: the Case of Entity Overshadowing
Vera Provatorova, Samarth Bhargav, Svitlana Vakulenko, Evangelos Kanoulas
摘要
Entity disambiguation (ED) is the last step of entity linking (EL), when candidate entities are reranked according to the context they appear in. All datasets for training and evaluating models for EL consist of convenience samples, such as news articles and tweets, that propagate the prior probability bias of the entity distribution towards more frequently occurring entities. It was previously shown that performance of EL systems on such datasets is overestimated, since it is possible to obtain higher accuracy scores by merely learning the prior. To provide a more adequate evaluation benchmark, we introduce the ShadowLink dataset, which includes 16K short text snippets annotated with entity mentions. We evaluate and report the performance of several popular EL systems on the ShadowLink benchmark. The results show a considerable difference in accuracy between common and uncommon ambiguous entities that require disambiguation, for all of the EL systems under evaluation, demonstrating the effects of prior probability bias and entity overshadowing.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper1
相关 Paper
- Evaluating Entity Disambiguation and the Role of Popularity in Retrieval-Based NLPAnthony Chen, Pallavi Gudipati, Shayne Longpre, Xiao Ling 等ACL 2021
- RAED: Retrieval-Augmented Entity Description Generation for Emerging Entity Linking and DisambiguationKarim Ghonim, Pere-Lluís Huguet Cabot, Riccardo Orlando, Roberto NavigliEMNLP 2025
- Fine-Grained Entity Typing for Domain Independent Entity LinkingYasumasa Onoe, Greg DurrettAAAI 2020 · 被引用 94 次
- Entity Disambiguation with Extreme Multi-label RankingJyun-Yu Jiang, Wei-Cheng Chang, Jiong Zhang, Cho-Jui Hsieh 等WWW 2024 · 被引用 6 次
- EDIN: An End-to-end Benchmark and Pipeline for Unknown Entity Discovery and IndexingNora Kassner, Fabio Petroni, Mikhail Plekhanov, Sebastian Riedel 等EMNLP 2022 · 被引用 6 次
