Improving Zero-shot LLM Re-Ranker with Risk Minimization
Xiaowei Yuan, Zhao Yang, Yequan Wang, Jun Zhao, Kang Liu
Abstract
In the Retrieval-Augmented Generation (RAG) system, advanced Large Language Models (LLMs) have emerged as effective Query Likelihood Models (QLMs) in an unsupervised way, which re-rank documents based on the probability of generating the query given the content of a document. However, directly prompting LLMs to approximate QLMs inherently is biased, where the estimated distribution might diverge from the actual document-specific distribution. In this study, we introduce a novel framework, , which leverages Bayesian decision theory to both quantify and mitigate this estimation bias. Specifically, reformulates the problem as maximizing the probability of document generation, thereby harmonizing the optimization of query and document generation probabilities under a unified risk minimization objective. Our empirical results indicate that significantly enhances re-ranking, particularly in improving the Top-1 accuracy. It benefits the QA tasks by achieving higher accuracy with fewer input documents.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a14a8dc3-d70b-4061-ba54-0205c8203bbaCited by top-tier papers5
- ARise: Towards Knowledge-Augmented Reasoning via Risk-Adaptive SearchYize Zhang, Tianshu Wang, Sirui Chen, Kun Wang et al.ACL 2025 · 6 citations
- PaSa: An LLM Agent for Comprehensive Academic Paper SearchYichen He, Guanhua Huang, Peiyuan Feng, Yuan Lin et al.ACL 2025
- Precise Zero-Shot Pointwise Ranking with LLMs through Post-Aggregated Global Context InformationKehan Long, Shasha Li, Chen Xu, Jintao Tang et al.SIGIR 2025
- Efficient Prior-Guided Reasoning for Robust Retrieval-Augmented Generation under ConflictsXiaowei Yuan, Ziyang Huang, Zhao Yang, Yequan Wang et al.ACL 2026
- Optimizing RAG Rerankers with LLM Feedback via Reinforcement LearningYuhang Wu, Xiangqing Shen, Fanfan Wang, Cangqi Zhou et al.ACL 2026
Builds on5
- Is ChatGPT Good at Search? Investigating Large Language Models as Re-Ranking AgentsWeiwei Sun, Lingyong Yan, Xinyu Ma, Shuaiqiang Wang et al.EMNLP 2023 · 182 citations
- Dense Passage Retrieval for Open-Domain Question AnsweringVladimir Karpukhin, Barlas Oguz, Sewon Min, Patrick Lewis et al.EMNLP 2020 · 142 citations
- Training Data is More Valuable than You Think: A Simple and Effective Method by Retrieving from Training DataShuohang Wang, Yichong Xu, Yuwei Fang, Yang Liu et al.ACL 2022 · 115 citations
- Improving Passage Retrieval with Zero-Shot Question GenerationDevendra Singh Sachan, Mike Lewis, Mandar Joshi, Armen Aghajanyan et al.EMNLP 2022 · 69 citations
- End-to-End Training of Neural Retrievers for Open-Domain Question AnsweringDevendra Singh Sachan, Mostofa Patwary, Mohammad Shoeybi, Neel Kant et al.ACL 2021
Related papers
- s3: You Don't Need That Much Data to Train a Search Agent via RLPengcheng Jiang, Xueqiang Xu, Jiacheng Lin, Jinfeng Xiao et al.EMNLP 2025
- REALM: Recursive Relevance Modeling for LLM-based Document Re-RankingPinhuan Wang, Zhiqiu Xia, Chunhua Liao, Feiyi Wang et al.EMNLP 2025
- Reliable Decision‑Making via Calibration‑Oriented Retrieval‑Augmented GenerationChaeyun Jang, Deukhwan Cho, Seanie Lee, Hyungi Lee et al.NeurIPS 2025 · 5 citations
- UniRAG: Unified Query Understanding Method for Retrieval Augmented GenerationRui Li, Liyang He, Qi Liu, Zheng Zhang et al.ACL 2025
- DynamicRAG: Leveraging Outputs of Large Language Model as Feedback for Dynamic Reranking in Retrieval-Augmented GenerationJiashuo Sun, Xianrui Zhong, Sizhe Zhou, Jiawei HanNeurIPS 2025 · 19 citations
