A Generalised and Adaptable Reinforcement Learning Stopping Method
Reem Bin Hezam, Mark Stevenson
摘要
This paper presents a Technology Assisted Review (TAR) stopping approach based on Reinforcement Learning (RL). Previous such approaches offered limited control over stopping behaviour, such as fixing the target recall and tradeoff between preferring to maximise recall or cost. These limitations are overcome by introducing a novel RL environment, GRLStop, that allows a single model to be applied to multiple target recalls, balances the recall/cost tradeoff and integrates a classifier. Experiments were carried out on six benchmark datasets (CLEF e-Health datasets 2017-9, TREC Total Recall, TREC Legal and Reuters RCV1) at multiple target recall levels. Results showed that the proposed approach to be effective compared to multiple baselines in addition to offering greater flexibility.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper3
- Self-Supervised Reinforcement Learning for Recommender SystemsXin Xin, Alexandros Karatzoglou, Ioannis Arapakis, Joemon M. JoseSIGIR 2020 · 被引用 217 次
- A Reinforcement Learning Framework for Relevance FeedbackAli Montazeralghaem, Hamed Zamani, James AllanSIGIR 2020 · 被引用 38 次
- Contrastive State Augmentations for Reinforcement Learning-Based Recommender SystemsZhaochun Ren, Na Huang, Yidan Wang, Pengjie Ren 等SIGIR 2023 · 被引用 20 次
相关 Paper
- Decision-Theoretic Stopping Rules for Document ScreeningAaron H. A. Fletcher, Mark StevensonSIGIR 2026
- Improving Multi-party Dialogue Generation via Topic and Rhetorical CoherenceYaxin Fan, Peifeng Li, Qiaoming ZhuEMNLP 2024 · 被引用 1 次
- TGRL: An Algorithm for Teacher Guided Reinforcement LearningIdan Shenfeld, Zhang-Wei Hong, Aviv Tamar, Pulkit AgrawalICML 2023 · 被引用 22 次
- Learning to Truncate Ranked Lists for Information RetrievalChen Wu, Ruqing Zhang, Jiafeng Guo, Yixing Fan 等AAAI 2021 · 被引用 10 次
- Extracting Relevant Information from User's Utterances in Conversational Search and RecommendationAli Montazeralghaem, James AllanKDD 2022 · 被引用 5 次
