Preference Learning of Latent Decision Utilities with a Human-like Model of Preferential Choice
Sebastiaan De Peuter, Shibei Zhu, Yujia Guo, Andrew Howes, Samuel Kaski
摘要
Preference learning methods make use of models of human choice in order to infer the latent utilities that underlie human behavior. However, accurate modeling of human choice behavior is challenging due to a range of context effects that arise from how humans contrast and evaluate options. Cognitive science has proposed several models that capture these intricacies but, due to their intractable nature, work on preference learning has, in practice, had to rely on tractable but simplified variants of the well-known Bradley-Terry model. In this paper, we take one state-of-the-art intractable cognitive model and propose a tractable surrogate that is suitable for deployment in preference learning. We then introduce a mechanism for fitting the surrogate to human data and extend it to account for data that cannot be explained by the original cognitive model. We demonstrate on large-scale human data that this model produces significantly better inferences on static and actively elicited data than existing Bradley-Terry variants. We further show in simulation that when using this model for preference learning, we can significantly improve utility in a range of real-world tasks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- A tale of two tails: Preferred and anti-preferred natural stimuli in visual cortexRabia Gondur, Patricia L. Stan, Matthew A. Smith, Benjamin R. CowleyICLR 2026 · 被引用 3 次
- Risk-aware Direct Preference Optimization under Nested Risk MeasureLijun Zhang, Lin Li, Yajie Qi, Huizhong Song 等NeurIPS 2025 · 被引用 3 次
它引用的顶会 Paper5
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida 等NeurIPS 2022 · 被引用 24,707 次
- Learning to summarize with human feedbackNisan Stiennon, Long Ouyang, Jeffrey Wu, Daniel M. Ziegler 等NeurIPS 2020 · 被引用 124 次
- Aligning Synthetic Medical Images with Clinical Knowledge using Human FeedbackShenghuan Sun, Gregory M. Goldgof, Atul J. Butte, Ahmed M. AlaaNeurIPS 2023 · 被引用 27 次
- Preference Modeling with Context-Dependent Salient FeaturesAmanda Bower, Laura BalzanoICML 2020 · 被引用 16 次
- Learning Interpretable Feature Context Effects in Discrete ChoiceKiran Tomlinson, Austin R. BensonKDD 2021 · 被引用 2 次
相关 Paper
- What Does Preference Learning Recover from Pairwise Comparison Data?Rattana Pukdee, Nina Balcan, Pradeep RavikumarICML 2026 · 被引用 1 次
- Generalizing while preserving monotonicity in comparison-based preference learning modelsJulien Fageot, Peva Blanchard, Gilles Bareilles, Lê-Nguyên HoangNeurIPS 2025 · 被引用 2 次
- Towards Cognitively-Faithful Decision-Making Models to Improve AI AlignmentCyrus Cousins, Vijay Keswani, Vincent Conitzer, Hoda Heidari 等ICLR 2026 · 被引用 2 次
- Online Compatible Reward Identification from Preference FeedbackSimone Drago, Marco Mussi, Alberto Maria MetelliICML 2026
- Act-Adaptive Margin: Dynamically Calibrating Reward Models for Subjective AmbiguityFeiteng Fang, Dingwei Chen, Xiang Huang, Ting-En Lin 等ACL 2026 · 被引用 3 次
