Breaking the Lens of the Telescope: Online Relevance Estimation over Large Retrieval Sets
Mandeep Rathee, Venktesh V, Sean MacAvaney, Avishek Anand
摘要
Advanced relevance models, such as those that use large language models (LLMs), provide highly accurate relevance estimations.However, their computational costs make them infeasible for processing large document corpora.To address this, retrieval systems often employ a telescoping approach, where computationally efficient but less precise lexical and semantic retrievers filter potential candidates for further ranking.However, this approach heavily depends on the quality of early-stage retrieval, which can potentially exclude relevant documents early in the process.In this work, we propose a novel paradigm for re-ranking called online relevance estimation that continuously updates relevance estimates for a query throughout the ranking process.Instead of re-ranking a fixed set of top-k documents in a single step, online relevance estimation iteratively re-scores smaller subsets of the most promising documents while adjusting relevance scores for the remaining pool based on the estimations from the final model using an online bandit-based algorithm.This dynamic process mitigates the recall limitations of telescoping systems by re-prioritizing documents initially deemed less relevant by earlier stages-including those completely excluded by earlier-stage retrievers.We validate our approach on TREC benchmarks under two scenarios: hybrid retrieval and adaptive retrieval.Experimental results demonstrate that our method is sample-efficient and significantly improves recall, highlighting the effectiveness of our online relevance estimation framework for modern search systems.https://github.com/elixir
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- When More Reformulations Hurt: Avoiding Drift using Ranker FeedbackVenktesh V, Mandeep Rathee, Avishek AnandSIGIR 2026 · 被引用 1 次
- Sample Efficient Demonstration Selection for In-Context LearningKiran Purohit, Venktesh V, Sourangshu Bhattacharya, Avishek AnandICML 2025
- PLAID-PRF: Pseudo-Relevance Feedback with Centroid-like Tokens in PLAIDXiao Wang, Sean MacAvaney, Craig MacdonaldSIGIR 2026
它引用的顶会 Paper5
- Efficiently Teaching an Effective Dense Retriever with Balanced Topic Aware SamplingSebastian Hofstätter, Sheng-Chieh Lin, Jheng-Hong Yang, Jimmy Lin 等SIGIR 2021 · 被引用 297 次
- Is ChatGPT Good at Search? Investigating Large Language Models as Re-Ranking AgentsWeiwei Sun, Lingyong Yan, Xinyu Ma, Shuaiqiang Wang 等EMNLP 2023 · 被引用 182 次
- Dense Passage Retrieval for Open-Domain Question AnsweringVladimir Karpukhin, Barlas Oguz, Sewon Min, Patrick Lewis 等EMNLP 2020 · 被引用 142 次
- Lexically-Accelerated Dense RetrievalHrishikesh Kulkarni, Sean MacAvaney, Nazli Goharian, Ophir FriederSIGIR 2023 · 被引用 30 次
- Efficient Neural Ranking using Forward IndexesJurek Leonhardt, Koustav Rudra, Megha Khosla, Abhijit Anand 等WWW 2022 · 被引用 16 次
相关 Paper
- AcuRank: Uncertainty-Aware Adaptive Computation for Listwise RerankingSoyoung Yoon, Gyuwan Kim, Gyu-Hwung Cho, Seung-won HwangNeurIPS 2025 · 被引用 15 次
- REALM: Recursive Relevance Modeling for LLM-based Document Re-RankingPinhuan Wang, Zhiqiu Xia, Chunhua Liao, Feiyi Wang 等EMNLP 2025
- Contextual Relevance and Adaptive Sampling for LLM-Based Document RerankingJerry Huang, Siddarth Madala, Cheng Niu, Julia Hockenmaier 等ACL 2026 · 被引用 3 次
- Attention in Large Language Models Yields Efficient Zero-Shot Re-RankersShijie Chen, Bernal Jimenez Gutierrez, Yu SuICLR 2025
- Self-Calibrated Listwise Reranking with Large Language ModelsRuiyang Ren, Yuhao Wang, Kun Zhou, Wayne Xin Zhao 等WWW 2025 · 被引用 12 次
