Decision Aggregation under Quantal Response
Zhihuan Huang, Yichong Xia, Yuqing Kong
摘要
The effectiveness of collective decision-making is often challenged by the bounded rationality and inherent stochasticity of individual agents. We investigate this by analyzing how to aggregate decisions from experts, each receiving a private signal about an unknown state. Assuming signals are conditionally independent and identically distributed, we depart from the fully rational paradigm and model expert behavior using quantal response—a stochastic choice model capturing bounded rationality. Within a minimax regret framework, we show that majority voting is the optimal robust aggregator when individual rationality falls below a certain threshold. Interestingly, such groups can outperform perfectly rational agents, as their decision randomness encodes weak but informative signals lost in deterministic behavior. We validate these findings using large language models (LLMs), which naturally exhibit quantal response via their temperature parameter. Aggregating moderately stochastic LLM outputs significantly improves accuracy on complex reasoning tasks, highlighting bounded rationality not as a limitation, but as a potential strength in collective intelligence.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper4
- LLM Strategic Reasoning: Agentic Study through Behavioral Game TheoryJingru Jia, Zehua Yuan, Junhao Pan, Paul McNamara 等NeurIPS 2025 · 被引用 23 次
- Robust Decision Aggregation with Second-order InformationYuqi Pan, Zhaohua Chen, Yuqing KongWWW 2024 · 被引用 9 次
- Rationality-Robust Information Design: Bayesian Persuasion under Quantal ResponseYiding Feng, Chien-Ju Ho, Wei TangSODA 2024 · 被引用 3 次
- Robust Aggregation with Adversarial ExpertsYongkang Guo, Yuqing KongWWW 2025 · 被引用 2 次
相关 Paper
- Beyond Majority Voting: LLM Aggregation by Leveraging Higher-Order InformationRui Ai, Yuqi Pan, David Simchi-Levi, Milind Tambe 等ICML 2026 · 被引用 20 次
- Online Mixture of Experts: No-Regret Learning for Optimal Collective Decision-MakingLarkin Liu, Jalal EtesamiNeurIPS 2025 · 被引用 2 次
- Social Dynamics as Critical Vulnerabilities that Undermine Objective Decision-Making in LLM CollectivesChanggeon Ko, Jisu Shin, Hoyun Song, Huije Lee 等ACL 2026 · 被引用 1 次
- Multi-Agent Debate for LLM Judges with Adaptive Stability DetectionTianyu Hu, Zhen Tan, Song Wang, Huaizhi Qu 等NeurIPS 2025 · 被引用 25 次
- Best-of-Infinity: Asymptotic Performance of Test-Time LLM EnsemblingJunpei Komiyama, Daisuke Oba, Masafumi OyamadaICLR 2026 · 被引用 2 次
