Preference Learning of Latent Decision Utilities with a Human-like Model of Preferential Choice
Sebastiaan De Peuter, Shibei Zhu, Yujia Guo, Andrew Howes, Samuel Kaski
Abstract
Preference learning methods make use of models of human choice in order to infer the latent utilities that underlie human behavior. However, accurate modeling of human choice behavior is challenging due to a range of context effects that arise from how humans contrast and evaluate options. Cognitive science has proposed several models that capture these intricacies but, due to their intractable nature, work on preference learning has, in practice, had to rely on tractable but simplified variants of the well-known Bradley-Terry model. In this paper, we take one state-of-the-art intractable cognitive model and propose a tractable surrogate that is suitable for deployment in preference learning. We then introduce a mechanism for fitting the surrogate to human data and extend it to account for data that cannot be explained by the original cognitive model. We demonstrate on large-scale human data that this model produces significantly better inferences on static and actively elicited data than existing Bradley-Terry variants. We further show in simulation that when using this model for preference learning, we can significantly improve utility in a range of real-world tasks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4cfdb4d1-2a07-4658-9d32-b5d0fc327837Cited by top-tier papers2
- A tale of two tails: Preferred and anti-preferred natural stimuli in visual cortexRabia Gondur, Patricia L. Stan, Matthew A. Smith, Benjamin R. CowleyICLR 2026 · 3 citations
- Risk-aware Direct Preference Optimization under Nested Risk MeasureLijun Zhang, Lin Li, Yajie Qi, Huizhong Song et al.NeurIPS 2025 · 3 citations
Builds on5
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida et al.NeurIPS 2022 · 24,707 citations
- Learning to summarize with human feedbackNisan Stiennon, Long Ouyang, Jeffrey Wu, Daniel M. Ziegler et al.NeurIPS 2020 · 124 citations
- Aligning Synthetic Medical Images with Clinical Knowledge using Human FeedbackShenghuan Sun, Gregory M. Goldgof, Atul J. Butte, Ahmed M. AlaaNeurIPS 2023 · 27 citations
- Preference Modeling with Context-Dependent Salient FeaturesAmanda Bower, Laura BalzanoICML 2020 · 16 citations
- Learning Interpretable Feature Context Effects in Discrete ChoiceKiran Tomlinson, Austin R. BensonKDD 2021 · 2 citations
Related papers
- What Does Preference Learning Recover from Pairwise Comparison Data?Rattana Pukdee, Nina Balcan, Pradeep RavikumarICML 2026 · 1 citation
- Generalizing while preserving monotonicity in comparison-based preference learning modelsJulien Fageot, Peva Blanchard, Gilles Bareilles, Lê-Nguyên HoangNeurIPS 2025 · 2 citations
- Towards Cognitively-Faithful Decision-Making Models to Improve AI AlignmentCyrus Cousins, Vijay Keswani, Vincent Conitzer, Hoda Heidari et al.ICLR 2026 · 2 citations
- Online Compatible Reward Identification from Preference FeedbackSimone Drago, Marco Mussi, Alberto Maria MetelliICML 2026
- Act-Adaptive Margin: Dynamically Calibrating Reward Models for Subjective AmbiguityFeiteng Fang, Dingwei Chen, Xiang Huang, Ting-En Lin et al.ACL 2026 · 3 citations
