A Reinforcement Learning Framework for Relevance Feedback
Ali Montazeralghaem, Hamed Zamani, James Allan
Abstract
We present RML, the first known general reinforcement learning framework for relevance feedback that directly optimizes any desired retrieval metric, including precision-oriented, recall-oriented, and even diversity metrics: RML can be easily extended to directly optimize any arbitrary user satisfaction signal. Using the RML framework, we can select effective feedback terms and weight them appropriately, improving on past methods that fit parameters to feedback algorithms using heuristic approaches or methods that do not directly optimize for retrieval performance. Learning an effective relevance feedback model is not trivial since the true feedback distribution is unknown. Experiments on standard TREC collections compare RML to existing feedback algorithms, demonstrate the effectiveness of RML at optimizing for MAP and α-n DCG, and show the impact on related measures.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9c0c2409-d711-402d-9675-eee85e720e48Cited by top-tier papers5
- User Retention-oriented Recommendation with Decision TransformerKesen Zhao, Lixin Zou, Xiangyu Zhao, Maolin Wang et al.WWW 2023 · 38 citations
- LoL: A Comparative Regularization Loss over Query Reformulation Losses for Pseudo-Relevance FeedbackYunchang Zhu, Liang Pang, Yanyan Lan, Huawei Shen et al.SIGIR 2022 · 6 citations
- Extracting Relevant Information from User's Utterances in Conversational Search and RecommendationAli Montazeralghaem, James AllanKDD 2022 · 5 citations
- A Generalised and Adaptable Reinforcement Learning Stopping MethodReem Bin Hezam, Mark StevensonSIGIR 2025 · 1 citation
- Algorithmic Vibe in Information RetrievalAli Montazeralghaem, Nick Craswell, Ryen W. White, Ahmed Hassan Awadallah et al.WWW 2023
Related papers
- A Reference-Dependent Model for Web Search Evaluation: Understanding and Measuring the Experience of Boundedly Rational UsersNuo Chen, Jiqun Liu, Tetsuya SakaiWWW 2023 · 21 citations
- New Insights into Metric Optimization for Ranking-based RecommendationRoger Zhe Li, Julián Urbano, Alan HanjalicSIGIR 2021 · 6 citations
- Optimization Methods for Personalizing Large Language Models through Retrieval AugmentationAlireza Salemi, Surya Kallumadi, Hamed ZamaniSIGIR 2024 · 52 citations
- Enhancing Generative Retrieval with Reinforcement Learning from Relevance FeedbackYujia Zhou, Zhicheng Dou, Ji-Rong WenEMNLP 2023 · 14 citations
- Unsupervised Large Language Model Alignment for Information Retrieval via Contrastive FeedbackQian Dong, Yiding Liu, Qingyao Ai, Zhijing Wu et al.SIGIR 2024 · 9 citations
