Carrot and Stick: Eliciting Comparison Data and Beyond
Yiling Chen, Shi Feng, Fang-Yi Yu
摘要
Comparison data elicited from people are fundamental to many machine learning tasks, including reinforcement learning from human feedback for large language models and estimating ranking models. They are typically subjective and not directly verifiable. How to truthfully elicit such comparison data from rational individuals? We design peer prediction mechanisms for eliciting comparison data using a bonus-penalty payment. Our design leverages on the strong stochastic transitivity for comparison data to create symmetrically strongly truthful mechanisms such that truth-telling 1) forms a strict Bayesian Nash equilibrium, and 2) yields the highest payment among all symmetric equilibria. Each individual only needs to evaluate one pair of items and report her comparison in our mechanism. We further extend the bonus-penalty payment concept to eliciting networked data, designing a symmetrically strongly truthful mechanism when agents' private signals are sampled according to the Ising models. We provide the necessary and sufficient conditions for our bonus-penalty payment to have truth-telling as a strict Bayesian Nash equilibrium. Experiments on two real-world datasets further support our theoretical discoveries.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper5
- Chatbot Arena: An Open Platform for Evaluating LLMs by Human PreferenceWei-Lin Chiang, Lianmin Zheng, Ying Sheng, Anastasios Nikolas Angelopoulos 等ICML 2024 · 被引用 1,212 次
- Dominantly Truthful Multi-task Peer Prediction with a Constant Number of TasksYuqing KongSODA 2020 · 被引用 33 次
- Information Elicitation from Rowdy CrowdsGrant Schoenebeck, Fang-Yi Yu, Yichi ZhangWWW 2021 · 被引用 18 次
- High-Effort Crowds: Limited Liability via TournamentsYichi Zhang, Grant SchoenebeckWWW 2023 · 被引用 10 次
- Peer Prediction for Learning AgentsShi Feng, Fang-Yi Yu, Yiling ChenNeurIPS 2022 · 被引用 9 次
相关 Paper
- Stochastically Dominant Peer PredictionYichi Zhang, Shengwei Xu, Grant Schoenebeck, David M. PennockNeurIPS 2025 · 被引用 2 次
- Truthful Data Acquisition via Peer PredictionYiling Chen, Yiheng Shen, Shuran ZhengNeurIPS 2020 · 被引用 35 次
- Multitask Peer Prediction With Task-dependent StrategiesYichi Zhang, Grant SchoenebeckWWW 2023 · 被引用 7 次
- Incentivizing Truthful Language Models via Peer Elicitation GamesBaiting Chen, Tong Zhu, Jiale Han, Lexin Li 等NeurIPS 2025 · 被引用 9 次
- Peer Neighborhood Mechanisms: A Framework for Mechanism GeneralizationAdam Richardson, Boi FaltingsAAAI 2024 · 被引用 1 次
