Gumbel Reranking: Differentiable End-to-End Reranker Optimization
Siyuan Huang, Zhiyuan Ma, Jintao Du, Changhua Meng, Weiqiang Wang, Jingwen Leng, Minyi Guo, Zhouhan Lin
Abstract
RAG systems rely on rerankers to identify relevant documents. However, fine-tuning these models remains challenging due to the scarcity of annotated query-document pairs. Existing distillation-based approaches suffer from training-inference misalignment and fail to capture interdependencies among candidate documents. To overcome these limitations, we reframe the reranking process as an attentionmask problem and propose Gumbel Reranking, an end-to-end training framework for rerankers aimed at minimizing the training-inference gap. In our approach, reranker optimization is reformulated as learning a stochastic, documentwise Top-k attention mask using the Gumbel trick and relaxed top-k sampling. This formulation enables end-to-end optimization by minimizing the overall language loss. Experiments across various settings consistently demonstrate performance gains, including a 10.4% improvement in recall on HotpotQA for distinguishing indirectly relevant documents.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3ebd21aa-7a27-445b-b637-6a75dadae445Builds on17
- Retrieval-Augmented Generation for Knowledge-Intensive NLP TasksPatrick Lewis, Ethan Perez, Aleksandra Piktus, Fabio Petroni et al.NeurIPS 2020 · 19,162 citations
- Retrieval Augmented Language Model Pre-TrainingKelvin Guu, Kenton Lee, Zora Tung, Panupong Pasupat et al.ICML 2020 · 2,937 citations
- Improving Language Models by Retrieving from Trillions of TokensSebastian Borgeaud, Arthur Mensch, Jordan Hoffmann, Trevor Cai et al.ICML 2022 · 1,629 citations
- ColBERT: Efficient and Effective Passage Search via Contextualized Late Interaction over BERTOmar Khattab, Matei ZahariaSIGIR 2020 · 1,246 citations
- Distilling Knowledge from Reader to Retriever for Question AnsweringGautier Izacard, Edouard GraveICLR 2021 · 317 citations
Related papers
- GRO-RAG: Gradient-aware Re-rank Optimization for Multi-source Retrieval-Augmented GenerationSiyuan Chen, Hang Ding, Xiaoyu Kang, Jiechao GaoICLR 2026
- D-RAG: Differentiable Retrieval-Augmented Generation for Knowledge Graph Question AnsweringGuangze Gao, Zixuan Li, Chunfeng Yuan, Jiawei Li et al.EMNLP 2025
- Diversifying Differentiable Graph Retrieval with Topic-Adaptive Multi-Intent LearningDongcheon Lee, Ji-Yeon Park, Hye-Yoon Baek, Jimyeung Seo et al.WWW 2026
- In defense of dual-encoders for neural rankingAditya Krishna Menon, Sadeep Jayasumana, Ankit Singh Rawat, Seungyeon Kim et al.ICML 2022 · 29 citations
- Improving Zero-shot LLM Re-Ranker with Risk MinimizationXiaowei Yuan, Zhao Yang, Yequan Wang, Jun Zhao et al.EMNLP 2024 · 3 citations
