Ranking Interruptus: When Truncated Rankings Are Better and How to Measure That
Enrique Amigó, Stefano Mizzaro, Damiano Spina
Abstract
Most of information retrieval effectiveness evaluation metrics assume that systems appending irrelevant documents at the bottom of the ranking are as effective as (or not worse than) systems that have a stopping criteria to 'truncate' the ranking at the right position to avoid retrieving those irrelevant documents at the end. It can be argued, however, that such truncated rankings are more useful to the end user. It is thus important to understand how to measure retrieval effectiveness in this scenario. In this paper we provide both theoretical and experimental contributions. We first define formal properties to analyze how effectiveness metrics behave when evaluating truncated rankings. Our theoretical analysis shows that de-facto standard metrics do not satisfy desirable properties to evaluate truncated rankings: only Observational Information Effectiveness (OIE) -- a metric based on Shannon's information theory -- satisfies them all. We then perform experiments to compare several metrics on nine TREC datasets. According to our experimental results, the most appropriate metrics for truncated rankings are OIE and a novel extension of Rank-Biased Precision that adds a user effort factor penalizing the retrieval of irrelevant documents.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d76dbb94-8c14-46dc-896b-0596ca1e2eb3Cited by top-tier papers1
Ask how each one uses itBuilds on2
- An Effectiveness Metric for Ordinal Classification: Formal Properties and Experimental ResultsEnrique Amigó, Julio Gonzalo, Stefano Mizzaro, Jorge Carrillo-de-AlbornozACL 2020 · 39 citations
- Evaluation Measures Based on Preference GraphsCharles L. A. Clarke, Chengxi Luo, Mark D. SmuckerSIGIR 2021 · 5 citations
Related papers
- A Flexible Framework for Offline Effectiveness MetricsAlistair Moffat, Joel Mackenzie, Paul Thomas, Leif AzzopardiSIGIR 2022 · 41 citations
- The Treatment of Ties in Rank-Biased OverlapMatteo Corsi, Julián UrbanoSIGIR 2024 · 11 citations
- A Reference-Dependent Model for Web Search Evaluation: Understanding and Measuring the Experience of Boundedly Rational UsersNuo Chen, Jiqun Liu, Tetsuya SakaiWWW 2023 · 21 citations
- MileCut: A Multi-view Truncation Framework for Legal Case RetrievalFuda Ye, Shuangyin LiWWW 2024 · 2 citations
- Incorporating Retrieval Information into the Truncation of Ranking Lists for Better Legal SearchYixiao Ma, Qingyao Ai, Yueyue Wu, Yunqiu Shao et al.SIGIR 2022 · 20 citations
