ReasonRank: Empowering Passage Ranking with Strong Reasoning Ability
Wenhan Liu, Xinyu Ma, Weiwei Sun, Yutao Zhu, Yuchen Li, Dawei Yin, Zhicheng Dou
摘要
Large Language Model (LLM) based listwise ranking has shown superior performance in many passage ranking tasks. With the development of Large Reasoning Models (LRMs), many studies have demonstrated that step-by-step reasoning during test-time helps improve listwise ranking performance. However, due to the scarcity of reasoning-intensive training data, existing rerankers perform poorly in many complex ranking scenarios, and the ranking ability of reasoning-intensive rerankers remains largely underdeveloped. In this paper, we first propose an automated reasoning-intensive training data synthesis framework, which sources training queries and passages from diverse domains and applies DeepSeek-R1 to generate high-quality training labels. To empower the listwise reranker with strong reasoning ability, we further propose a two-stage training approach, which includes a cold-start supervised fine-tuning (SFT) stage and a reinforcement learning (RL) stage. During the RL stage, we design a novel multi-view ranking reward tailored to the multi-turn nature of listwise ranking. Extensive experiments demonstrate that our trained reasoning-intensive reranker ReasonRank outperforms existing baselines significantly and also achieves much lower latency than the pointwise reranker. Our codes are available at https://github.com/8421BCD/ReasonRank.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- AcuRank: Uncertainty-Aware Adaptive Computation for Listwise RerankingSoyoung Yoon, Gyuwan Kim, Gyu-Hwung Cho, Seung-won HwangNeurIPS 2025 · 被引用 15 次
- ReasonEmbed: Enhanced Text Embeddings for Reasoning-Intensive Document RetrievalJianlyu Chen, Junwei Lan, Chaofan Li, Defu Lian 等ACL 2026 · 被引用 13 次
- Rethinking Reasoning in Document Ranking: Why Chain-of-Thought Falls ShortXuan Lu, Haohang Huang, Rui Meng, Yaohui Jin 等ICLR 2026 · 被引用 11 次
- Retro*: Optimizing LLMs for Reasoning-Intensive Document RetrievalJunwei Lan, Jianlyu Chen, Zheng Liu, Chaofan Li 等ICLR 2026 · 被引用 8 次
- ZeroGR: A Generalizable and Scalable Framework for Zero-Shot Generative RetrievalWeiwei Sun, Keyi Kong, Xinyu Ma, Shuaiqiang Wang 等ICLR 2026 · 被引用 6 次
它引用的顶会 Paper11
- FlashAttention-2: Faster Attention with Better Parallelism and Work PartitioningTri DaoICLR 2024 · 被引用 2,600 次
- Self-Consistency Improves Chain of Thought Reasoning in Language ModelsXuezhi Wang, Jason Wei, Dale Schuurmans, Quoc V. Le 等ICLR 2023 · 被引用 681 次
- Is ChatGPT Good at Search? Investigating Large Language Models as Re-Ranking AgentsWeiwei Sun, Lingyong Yan, Xinyu Ma, Shuaiqiang Wang 等EMNLP 2023 · 被引用 182 次
- Improving Passage Retrieval with Zero-Shot Question GenerationDevendra Singh Sachan, Mike Lewis, Mandar Joshi, Armen Aghajanyan 等EMNLP 2022 · 被引用 69 次
- A Setwise Approach for Effective and Highly Efficient Zero-shot Ranking with Large Language ModelsShengyao Zhuang, Honglei Zhuang, Bevan Koopman, Guido ZucconSIGIR 2024 · 被引用 60 次
相关 Paper
- ERank: Fusing Supervised Fine-Tuning and Reinforcement Learning for Effective and Efficient Text RerankingYuzheng Cai, Yanzhao Zhang, Dingkun Long, Mingxin Li 等AAAI 2026 · 被引用 6 次
- REARANK: Reasoning Re-ranking Agent via Reinforcement LearningLe Zhang, Bo Wang, Xipeng Qiu, Siva Reddy 等EMNLP 2025 · 被引用 14 次
- ReSearch: Learning to Reason with Search for LLMs via Reinforcement LearningMingyang Chen, Linzhuang Sun, Tianpeng Li, Haoze Sun 等NeurIPS 2025 · 被引用 125 次
- Is PRM Necessary? Problem-Solving RL Implicitly Induces PRM Capability in LLMsZhangyin Feng, Qianglong Chen, Ning Lu, Yongqian Li 等NeurIPS 2025 · 被引用 16 次
- General-Reasoner: Advancing LLM Reasoning Across All DomainsXueguang Ma, Qian Liu, Dongfu Jiang, Ge Zhang 等NeurIPS 2025 · 被引用 153 次
