Peer Prediction for Learning Agents
Shi Feng, Fang-Yi Yu, Yiling Chen
摘要
Peer prediction refers to a collection of mechanisms for eliciting information from human agents when direct verification of the obtained information is unavailable. They are designed to have a game-theoretic equilibrium where everyone reveals their private information truthfully. This result holds under the assumption that agents are Bayesian and they each adopt a fixed strategy across all tasks. Human agents however are observed in many domains to exhibit learning behavior in sequential settings. In this paper, we explore the dynamics of sequential peer prediction mechanisms when participants are learning agents. We first show that the notion of no regret alone for the agents' learning algorithms cannot guarantee convergence to the truthful strategy. We then focus on a family of learning algorithms where strategy updates only depend on agents' cumulative rewards and prove that agents' strategies in the popular Correlated Agreement (CA) mechanism converge to truthful reporting when they use algorithms from this family. This family of algorithms is not necessarily no-regret, but includes several familiar no-regret learning algorithms (e.g multiplicative weight update and Follow the Perturbed Leader) as special cases. Simulation of several algorithms in this family as well as the -greedy algorithm, which is outside of this family, shows convergence to the truthful strategy in the CA mechanism. A fundamental challenge in many domains is to elicit high-quality information from people when directly verifying the acquired information is not feasible, either because the ground truth is not available or because it's too costly to obtain. Notable settings include asking people to label data for machine learning, having students perform peer grading in education, and soliciting customer feedback for products and services. The peer prediction literature has made impressive progress on this challenge in the past two decades, with many mechanisms that have desirable incentive properties developed for this problem [
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Nash Convergence of Mean-Based Learning Algorithms in First Price AuctionsXiaotie Deng, Xinyan Hu, Tao Lin, Weiqiang ZhengWWW 2022 · 被引用 16 次
- Equilibrium of Data Markets with ExternalitySafwan Hossain, Yiling ChenICML 2024 · 被引用 7 次
- Carrot and Stick: Eliciting Comparison Data and BeyondYiling Chen, Shi Feng, Fang-Yi YuNeurIPS 2024 · 被引用 5 次
- Incentive-Aligned Multi-Source LLM SummariesYanchen Jiang, Zhe Feng, Aranyak MehtaICLR 2026 · 被引用 2 次
- Truthfulness Despite Weak Supervision: Evaluating and Training LLMs Using Peer PredictionTianyi Qiu, Micah Carroll, Cameron AllenICLR 2026 · 被引用 1 次
它引用的顶会 Paper4
- Dominantly Truthful Multi-task Peer Prediction with a Constant Number of TasksYuqing KongSODA 2020 · 被引用 33 次
- Information Elicitation from Rowdy CrowdsGrant Schoenebeck, Fang-Yi Yu, Yichi ZhangWWW 2021 · 被引用 18 次
- Nash Convergence of Mean-Based Learning Algorithms in First Price AuctionsXiaotie Deng, Xinyan Hu, Tao Lin, Weiqiang ZhengWWW 2022 · 被引用 16 次
- Mechanisms for a No-Regret Agent: Beyond the Common PriorModibo K. Camara, Jason D. Hartline, Aleck C. JohnsenFOCS 2020 · 被引用 5 次
相关 Paper
- Stochastically Dominant Peer PredictionYichi Zhang, Shengwei Xu, Grant Schoenebeck, David M. PennockNeurIPS 2025 · 被引用 2 次
- Multitask Peer Prediction With Task-dependent StrategiesYichi Zhang, Grant SchoenebeckWWW 2023 · 被引用 7 次
- Information Elicitation Mechanisms for Statistical EstimationYuqing Kong, Grant Schoenebeck, Biaoshuai Tao, Fang-Yi YuAAAI 2020 · 被引用 22 次
- Truthful Data Acquisition via Peer PredictionYiling Chen, Yiheng Shen, Shuran ZhengNeurIPS 2020 · 被引用 35 次
- Convergence Analysis of No-Regret Bidding Algorithms in Repeated AuctionsZhe Feng, Guru Guruganesh, Christopher Liaw, Aranyak Mehta 等AAAI 2021 · 被引用 31 次
