Axioms for Learning from Pairwise Comparisons
Ritesh Noothigattu, Dominik Peters, Ariel D. Procaccia
摘要
To be well-behaved, systems that process preference data must satisfy certain conditions identified by economic decision theory and by social choice theory. In ML, preferences and rankings are commonly learned by fitting a probabilistic model to noisy preference data. The behavior of this learning process from the view of economic theory has previously been studied for the case where the data consists of rankings. In practice, it is more common to have only pairwise comparison data, and the formal properties of the associated learning problem are more challenging to analyze. We show that a large class of random utility models (including the Thurstone-Mosteller Model), when estimated using the MLE, satisfy a Pareto efficiency condition. These models also satisfy a strong monotonicity property, which implies that the learning process is responsive to input data. On the other hand, we show that these models fail certain other consistency conditions from social choice theory, and in particular do not always follow the majority opinion. Our results inform existing and future applications of random utility models for societal decision making.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- Axioms for AI Alignment from Human FeedbackLuise Ge, Daniel Halpern, Evi Micha, Ariel D. Procaccia 等NeurIPS 2024 · 被引用 64 次
- Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences?Paul Gölz, Nika Haghtalab, Kunhe YangNeurIPS 2025 · 被引用 29 次
- Barter Exchange with Shared Item ValuationsJuan Luque, Sharmila Duppala, John P. Dickerson, Aravind SrinivasanWWW 2024 · 被引用 2 次
- Generalizing while preserving monotonicity in comparison-based preference learning modelsJulien Fageot, Peva Blanchard, Gilles Bareilles, Lê-Nguyên HoangNeurIPS 2025 · 被引用 2 次
- Towards Cognitively-Faithful Decision-Making Models to Improve AI AlignmentCyrus Cousins, Vijay Keswani, Vincent Conitzer, Hoda Heidari 等ICLR 2026 · 被引用 2 次
相关 Paper
- Beyond RLHF and NLHF: Population-Proportional Alignment under an Axiomatic FrameworkKihyun Kim, Jiawei Zhang, Asuman Ozdaglar, Pablo A. ParriloICLR 2026 · 被引用 5 次
- What Does Preference Learning Recover from Pairwise Comparison Data?Rattana Pukdee, Nina Balcan, Pradeep RavikumarICML 2026 · 被引用 1 次
- Preference Elicitation as Average-Case SortingDominik Peters, Ariel D. ProcacciaAAAI 2021 · 被引用 3 次
- Learning Correlated Reward Models: Statistical Barriers and OpportunitiesYeshwanth Cherapanamjeri, Constantinos Costis Daskalakis, Gabriele Farina, Sobhan MohammadpourICLR 2026 · 被引用 2 次
- Preference Modeling with Context-Dependent Salient FeaturesAmanda Bower, Laura BalzanoICML 2020 · 被引用 16 次
