A Framework for Studying AI Agent Behavior: Evidence from Consumer Choice Experiments
Manuel Cherep, Chengtian Ma, Abigail Xu, Maya Shaked, Pattie Maes, Nikhil Singh
摘要
Environments built for people are increasingly operated by a new class of economic actors: LLM-powered software agents making decisions on our behalf. These decisions range from our purchases to travel plans to medical treatment selection. Current evaluations of these agents largely focus on task competence, but we argue for a deeper assessment: how these agents choose when faced with realistic decisions. We introduce ABxLab, a framework for systematically probing agentic choice through controlled manipulations of option attributes and persuasive cues. We apply this to a realistic web-based shopping environment, where we vary prices, ratings, and psychological nudges, all of which are factors long known to shape human choice. We find that agent decisions shift predictably and substantially in response, revealing that agents are strongly biased choosers even without being subject to the cognitive constraints that shape human biases. This susceptibility reveals both risk and opportunity: risk, because agentic consumers may inherit and amplify human biases; opportunity, because consumer choice provides a powerful testbed for a behavioral science of AI agents, just as it has for the study of human behavior. We release our framework as an open benchmark for rigorous, scalable evaluation of agent decision-making.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper9
- WebShop: Towards Scalable Real-World Web Interaction with Grounded Language AgentsShunyu Yao, Howard Chen, John Yang, Karthik NarasimhanNeurIPS 2022 · 被引用 1,477 次
- Language Models can Solve Computer TasksGeunwoo Kim, Pierre Baldi, Stephen McAleerNeurIPS 2023 · 被引用 539 次
- Attacking Vision-Language Computer Agents via Pop-upsYanzhe Zhang, Tao Yu, Diyi YangACL 2025 · 被引用 99 次
- How do Large Language Models Navigate Conflicts between Honesty and Helpfulness?Ryan Liu, Theodore R. Sumers, Ishita Dasgupta, Thomas L. GriffithsICML 2024 · 被引用 33 次
- Do Large Language Models Perform the Way People Expect? Measuring the Human Generalization FunctionKeyon Vafa, Ashesh Rambachan, Sendhil MullainathanICML 2024 · 被引用 30 次
相关 Paper
- LLM Agents Can Be Choice-Supportive Biased Evaluators: An Empirical StudyNan Zhuang, Boyu Cao, Yi Yang, Jing Xu 等AAAI 2025 · 被引用 4 次
- The Role of Heuristics and Biases during Complex Choices with an AI TeammateNikolos Gurney, John H. Miller, David V. PynadathAAAI 2023 · 被引用 5 次
- STEER: Assessing the Economic Rationality of Large Language ModelsNarun Krishnamurthi Raman, Taylor Lundy, Samuel Joseph Amouyal, Yoav Levine 等ICML 2024 · 被引用 24 次
- For What It's Worth: Humans Overwrite Their Economic Self-interest to Avoid Bargaining With AI SystemsAlexander Erlei, Richeek Das, Lukas Meub, Avishek Anand 等CHI 2022 · 被引用 33 次
- Investigating the Impact of Dark Patterns on LLM-Based Web AgentsDevin Ersoy, Brandon Lee, Ananth Shreekumar, Arjun Arunasalam 等S&P 2026 · 被引用 16 次
