Ranking Interruptus: When Truncated Rankings Are Better and How to Measure That
Enrique Amigó, Stefano Mizzaro, Damiano Spina
摘要
Most of information retrieval effectiveness evaluation metrics assume that systems appending irrelevant documents at the bottom of the ranking are as effective as (or not worse than) systems that have a stopping criteria to 'truncate' the ranking at the right position to avoid retrieving those irrelevant documents at the end. It can be argued, however, that such truncated rankings are more useful to the end user. It is thus important to understand how to measure retrieval effectiveness in this scenario. In this paper we provide both theoretical and experimental contributions. We first define formal properties to analyze how effectiveness metrics behave when evaluating truncated rankings. Our theoretical analysis shows that de-facto standard metrics do not satisfy desirable properties to evaluate truncated rankings: only Observational Information Effectiveness (OIE) -- a metric based on Shannon's information theory -- satisfies them all. We then perform experiments to compare several metrics on nine TREC datasets. According to our experimental results, the most appropriate metrics for truncated rankings are OIE and a novel extension of Rank-Biased Precision that adds a user effort factor penalizing the retrieval of irrelevant documents.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper2
- An Effectiveness Metric for Ordinal Classification: Formal Properties and Experimental ResultsEnrique Amigó, Julio Gonzalo, Stefano Mizzaro, Jorge Carrillo-de-AlbornozACL 2020 · 被引用 39 次
- Evaluation Measures Based on Preference GraphsCharles L. A. Clarke, Chengxi Luo, Mark D. SmuckerSIGIR 2021 · 被引用 5 次
相关 Paper
- A Flexible Framework for Offline Effectiveness MetricsAlistair Moffat, Joel Mackenzie, Paul Thomas, Leif AzzopardiSIGIR 2022 · 被引用 41 次
- The Treatment of Ties in Rank-Biased OverlapMatteo Corsi, Julián UrbanoSIGIR 2024 · 被引用 11 次
- A Reference-Dependent Model for Web Search Evaluation: Understanding and Measuring the Experience of Boundedly Rational UsersNuo Chen, Jiqun Liu, Tetsuya SakaiWWW 2023 · 被引用 21 次
- MileCut: A Multi-view Truncation Framework for Legal Case RetrievalFuda Ye, Shuangyin LiWWW 2024 · 被引用 2 次
- Incorporating Retrieval Information into the Truncation of Ranking Lists for Better Legal SearchYixiao Ma, Qingyao Ai, Yueyue Wu, Yunqiu Shao 等SIGIR 2022 · 被引用 20 次
