LLM Agents Can Be Choice-Supportive Biased Evaluators: An Empirical Study
Nan Zhuang, Boyu Cao, Yi Yang, Jing Xu, Mingda Xu, Yuxiao Wang, Qi Liu
摘要
With Large Language Model (LLM) agents taking on more evaluation responsibilities in decision-making, it is essential to recognize their possible biases to guarantee fair and trustworthy AI-supported decisions. This study is the first to thoroughly examine the choice-supportive bias in LLM agents, a cognitive bias that is known to impact human decisionmaking and evaluation. We conduct experiments across 19 open/unopen-source LLM models in five scenarios at maximum, employing both memory-based and evaluation-based tasks adapted and redesigned from human cognitive studies. Our findings show that LLM agents may exhibit biased attribution or evaluation that supports their initial choices, and such bias may persist even if contextual hallucination is not observable. Key findings show that bias manifestation can differ greatly depending on prompt construction and context preservation, and the bias may be mitigated in larger models. Significantly, we observe that the bias increases when the agents perceive they are in control. Our extensive study involving 284 well-educated humans shows that, despite bias, certain LLM agents can still perform better than humans in similar evaluation tasks. This research contributes to the growing area of AI psychology, and the findings underscore the importance of addressing cognitive biases in LLM Agent systems, with wide-ranging implications spanning from improving AI-assisted decision-making to advancing AI safety and ethics. * These authors contributed equally. All equal-contributing authors worked jointly from the first day until the last day, sharing equal workload and contributions. N. Zhuang, B. Cao, M. Xu were responsible for designing and conducting the experiments, Y. Yang was responsible for coding, J. Xu performed the human study, Y. Wang and Q. Liu were responsible for reviewing and supervising.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper9
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Reflexion: language agents with verbal reinforcement learningNoah Shinn, Federico Cassano, Ashwin Gopinath, Karthik Narasimhan 等NeurIPS 2023 · 被引用 5,828 次
- Tree of Thoughts: Deliberate Problem Solving with Large Language ModelsShunyu Yao, Dian Yu, Jeffrey Zhao, Izhak Shafran 等NeurIPS 2023 · 被引用 5,068 次
- Buffer of Thoughts: Thought-Augmented Reasoning with Large Language ModelsLing Yang, Zhaochen Yu, Tianjun Zhang, Shiyi Cao 等NeurIPS 2024 · 被引用 144 次
- Large Language Models Play StarCraft II: Benchmarks and A Chain of Summarization ApproachWeiyu Ma, Qirui Mi, Yongcheng Zeng, Xue Yan 等NeurIPS 2024 · 被引用 122 次
相关 Paper
- Emulating Aggregate Human Choice Behavior and Biases with GPT Conversational AgentsStephen Pilli, Vivek NallurCHI 2026 · 被引用 2 次
- In Agents We Trust, but Who Do Agents Trust? Latent Source Preferences Steer LLM GenerationsMohammad Aflah Khan, Mahsa Amani, Soumi Das, Bishwamittra Ghosh 等ICLR 2026 · 被引用 4 次
- Exploring Prosocial Irrationality for LLM Agents: A Social Cognition ViewXuan Liu, Jie Zhang, Haoyang Shang, Song Guo 等ICLR 2025
- Cognitive Biases in LLM-Assisted Software DevelopmentXinyi Zhou, Zeinadsadat Saghi, Sadra Sabouri, Rahul Pandita 等ICSE 2026
- From Single to Societal: Analyzing Persona-Induced Bias in Multi-Agent InteractionsJiayi Li, Xiao Liu, Yansong FengAAAI 2026 · 被引用 3 次
