Logit Inflation in ListMLE: Theoretical Analysis and Mitigation Strategies
Riyaz Ahmad Bhat, Jaydeep Sen
2026Year
Abstract
Modern learning-to-rank methods often rely on listwise objectives that directly model and optimize relative document order over entire permutations. While these objectives improve ranking quality, they frequently produce models with highly inflated relevance scores whose magnitudes exceed what is necessary for meaningful document separation, leading to poor probabilistic calibration.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Related papers
- Learning to Rank with Variable Result Presentation LengthsNorman Knyazev, Harrie OosterhuisSIGIR 2025 · 1 citation
- Self-Calibrated Listwise Reranking with Large Language ModelsRuiyang Ren, Yuhao Wang, Kun Zhou, Wayne Xin Zhao et al.WWW 2025 · 12 citations
- Calibrated Preference Learning: The Case of Label RankingSanto Thies, Viktor Bengs, Timo Kaufmann, Sebastian Vollmer et al.ICML 2026
- FIRST: Faster Improved Listwise Reranking with Single Token DecodingRevanth Gangi Reddy, JaeHyeok Doo, Yifei Xu, Md. Arafat Sultan et al.EMNLP 2024 · 14 citations
- SetRank: Learning a Permutation-Invariant Ranking Model for Information RetrievalLiang Pang, Jun Xu, Qingyao Ai, Yanyan Lan et al.SIGIR 2020 · 113 citations
