Minimax Rate for Learning From Pairwise Comparisons in the BTL Model
Julien M. Hendrickx, Alex Olshevsky, Venkatesh Saligrama
摘要
We consider the problem of learning the qualities w 1 , . . . , w n of a collection of items by performing noisy comparisons among them. A standard assumption is that there is a fixed "comparison graph" and every neighboring pair of items is compared k times. We will study the popular Bradley-Terry-Luce model, where the probability that item i wins a comparison against j equals w i /(w i + w j ). The goal is to understand how the expected error in estimating the vector w = (w 1 , . . . , w n ) behaves in the regime when the number of comparisons k is large. Our contribution is the determination of the minimax rate up to a constant factor. We show that this rate is achieved by a simple algorithm based on weighted least squares, with weights determined from the empirical outcomes of the comparisons. This algorithm can be implemented in nearly linear time in the total number of comparisons.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Distributional Preference Learning: Understanding and Accounting for Hidden Context in RLHFAnand Siththaranjan, Cassidy Laidlaw, Dylan Hadfield-MenellICLR 2024 · 被引用 112 次
- Generalized Results for the Existence and Consistency of the MLE in the Bradley-Terry-Luce ModelHeejong Bong, Alessandro RinaldoICML 2022 · 被引用 23 次
- Inferring Dynamic Networks from Marginals with Iterative Proportional FittingSerina Chang, Frederic Koehler, Zhaonan Qu, Jure Leskovec 等ICML 2024 · 被引用 4 次
- Entrywise Error Bounds for Spectral Ranking with Semi-Random AdversariesDongmin Lee, Anuran Makur, Japneet SinghKDD 2026
- Energy-Based Preference Model Offers Better Offline Alignment than the Bradley-Terry Preference ModelYuzhong Hong, Hanshan Zhang, Junwei Bao, Hongfei Jiang 等ICML 2025
相关 Paper
- Rank Aggregation via Heterogeneous Thurstone Preference ModelsTao Jin, Pan Xu, Quanquan Gu, Farzad FarnoudAAAI 2020 · 被引用 19 次
- Estimation of Skill Distribution from a TournamentAli Jadbabaie, Anuran Makur, Devavrat ShahNeurIPS 2020 · 被引用 7 次
- Rank Aggregation from Pairwise Comparisons in the Presence of Adversarial CorruptionsArpit Agarwal, Shivani Agarwal, Sanjeev Khanna, Prathamesh PatilICML 2020 · 被引用 11 次
- The Sample Complexity of Best-k Items Selection from Pairwise ComparisonsWenbo Ren, Jia Liu, Ness B. ShroffICML 2020 · 被引用 14 次
- Score-Based Density Estimation from Pairwise ComparisonsPetrus Mikkola, Luigi Acerbi, Arto KlamiICLR 2026
