Generalized Results for the Existence and Consistency of the MLE in the Bradley-Terry-Luce Model
Heejong Bong, Alessandro Rinaldo
Abstract
Ranking problems based on pairwise comparisons, such as those arising in online gaming, often involve a large pool of items to order. In these situations, the gap in performance between any two items can be significant, and the smallest and largest winning probabilities can be very close to zero or one. Furthermore, each item may be compared only to a subset of all the items, so that not all pairwise comparisons are observed. In this paper, we study the performance of the Bradley-Terry-Luce model for ranking from pairwise comparison data under more realistic settings than those considered in the literature so far. In particular, we allow for near-degenerate winning probabilities and arbitrary comparison designs. We obtain novel results about the existence of the maximum likelihood estimator (MLE) and the corresponding estimation error without the bounded winning probability assumption commonly used in the literature and for arbitrary comparison graph topologies. Central to our approach is the reliance on the Fisher information matrix to express the dependence on the graph topologies and the impact of the values of the winning probabilities on the estimation risk and on the conditions for the existence of the MLE. Our bounds recover existing results as special cases but are more broadly applicable.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4733fb5a-ed12-4631-9d0d-307b36668b77Cited by top-tier papers10
- Direct Preference Optimization: Your Language Model is Secretly a Reward ModelRafael Rafailov, Archit Sharma, Eric Mitchell, Christopher D. Manning et al.NeurIPS 2023 · 10,924 citations
- GVPO: Group Variance Policy Optimization for Large Language Model Post-TrainingKaichen Zhang, Yuzhong Hong, Junwei Bao, Hongfei Jiang et al.NeurIPS 2025 · 35 citations
- An Analysis of Elo Rating Systems via Markov ChainsSam Olesker-Taylor, Luca ZanettiNeurIPS 2024 · 9 citations
- Inferring Dynamic Networks from Marginals with Iterative Proportional FittingSerina Chang, Frederic Koehler, Zhaonan Qu, Jure Leskovec et al.ICML 2024 · 4 citations
- Generative Bayesian Optimization: Generative Models as Acquisition FunctionsRafael Oliveira, Daniel M. Steinberg, Edwin V. BonillaICLR 2026 · 3 citations
Builds on1
Related papers
- Entrywise Error Bounds for Spectral Ranking with Semi-Random AdversariesDongmin Lee, Anuran Makur, Japneet SinghKDD 2026
- Rank Aggregation from Pairwise Comparisons in the Presence of Adversarial CorruptionsArpit Agarwal, Shivani Agarwal, Sanjeev Khanna, Prathamesh PatilICML 2020 · 11 citations
- Rank Aggregation via Heterogeneous Thurstone Preference ModelsTao Jin, Pan Xu, Quanquan Gu, Farzad FarnoudAAAI 2020 · 19 citations
- Generalized Bradley-Terry Models for Score Estimation from Paired ComparisonsJulien Fageot, Sadegh Farhadkhani, Lê-Nguyên Hoang, Oscar VillemaudAAAI 2024 · 12 citations
- Rethinking Reward Modeling in Preference-based Large Language Model AlignmentHao Sun, Yunyi Shen, Jean-Francois TonICLR 2025
