From Expectation to Evaluation: Expectation Cues Systematically Bias LLM and Human Judgment
Yiteng Sun, Danica Dillion, Kurt Gray, Mengtao Lyu, Zhuorui Zhang, Fan Li
Abstract
Expectation cues such as source labels, expertise signals, or identitybased indicators can bias how humans interpret and evaluate information. In high-stakes domains like healthcare, education, and law, such biases threaten the objectivity of decision-making. As LLMs increasingly provide decision support in these contexts, this study aims to examine whether LLMs exhibit expectation-driven bias akin to that of humans. Across two experiments (N = 1260), we manipulated expectations via priming statements and measured shifts in CHI '26, April 13-17, 2026, Barcelona, Spain Yiteng et al.
judgment scores. In both humans and LLMs, higher expectations led to more favorable evaluations for suggestions of equivalent quality, and greater mismatches between expectations and actual performance produced stronger judgment distortions. Notably, humans tended to adjust their evaluations unconsciously, whereas LLMs revised their outputs in a consistent and traceable manner. These findings reveal both shared sensitivities and distinct adjustment patterns, offering design insights for building expectation-aware AI systems that promote fair and transparent human-AI interaction.
• Human-centered computing → Empirical studies in HCI.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e8c1b4bc-da9f-41b8-8e44-9425d10530f0Builds on11
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma et al.NeurIPS 2022 · 22,562 citations
- To Trust or to Think: Cognitive Forcing Functions Can Reduce Overreliance on AI in AI-assisted Decision-makingZana Buçinca, Maja Barbara Malaya, Krzysztof Z. GajosCSCW 2021 · 962 citations
- Bias Runs Deep: Implicit Reasoning Biases in Persona-Assigned LLMsShashank Gupta, Vaishnavi Shrivastava, Ameet Deshpande, Ashwin Kalyan et al.ICLR 2024 · 212 citations
- Conceptual Metaphors Impact Perceptions of Human-AI CollaborationPranav Khadpe, Ranjay Krishna, Li Fei-Fei, Jeffrey T. Hancock et al.CSCW 2020 · 179 citations
- Explainable Active Learning (XAL): Toward AI Explanations as Interfaces for Machine TeachersBhavya Ghai, Q. Vera Liao, Yunfeng Zhang, Rachel K. E. Bellamy et al.CSCW 2020 · 107 citations
Related papers
- LLM Agents Can Be Choice-Supportive Biased Evaluators: An Empirical StudyNan Zhuang, Boyu Cao, Yi Yang, Jing Xu et al.AAAI 2025 · 4 citations
- Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-JudgeXin Sun, Di Wu, Sijing Qin, Isao Echizen et al.ACL 2026 · 2 citations
- Are Large Language Models Sensitive to the Motives Behind Communication?Addison J. Wu, Ryan Liu, Kerem Oktar, Theodore R. Sumers et al.NeurIPS 2025 · 9 citations
- Large Language Models Assume People are More Rational than We Really areRyan Liu, Jiayi Geng, Joshua C. Peterson, Ilia Sucholutsky et al.ICLR 2025
- Emulating Aggregate Human Choice Behavior and Biases with GPT Conversational AgentsStephen Pilli, Vivek NallurCHI 2026 · 2 citations
