Utilizing Human Behavior Modeling to Manipulate Explanations in AI-Assisted Decision Making: The Good, the Bad, and the Scary
Zhuoyan Li, Ming Yin
摘要
Recent advances in AI models have increased the integration of AI-based decision aids into the human decision making process. To fully unlock the potential of AI-assisted decision making, researchers have computationally modeled how humans incorporate AI recommendations into their final decisions, and utilized these models to improve human-AI team performance. Meanwhile, due to the ``black-box'' nature of AI models, providing AI explanations to human decision makers to help them rely on AI recommendations more appropriately has become a common practice. In this paper, we explore whether we can quantitatively model how humans integrate both AI recommendations and explanations into their decision process, and whether this quantitative understanding of human behavior from the learned model can be utilized to manipulate AI explanations, thereby nudging individuals towards making targeted decisions. Our extensive human experiments across various tasks demonstrate that human behavior can be easily influenced by these manipulated explanations towards targeted outcomes, regardless of the intent being adversarial or benign. Furthermore, individuals often fail to detect any anomalies in these explanations, despite their decisions being affected by them.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Plan-Then-Execute: An Empirical Study of User Trust and Team Performance When Using LLM Agents As A Daily AssistantGaole He, Gianluca Demartini, Ujwal GadirajuCHI 2025 · 被引用 91 次
- From Text to Trust: Empowering AI-assisted Decision Making with Adaptive LLM-powered AnalysisZhuoyan Li, Hangxiao Zhu, Zhuoran Lu, Ziang Xiao 等CHI 2025 · 被引用 30 次
- Human-LLM Collaborative Feature Engineering for Tabular DataZhuoyan Li, Aditya Bansal, Jinzhao Li, Shishuang He 等ICLR 2026 · 被引用 2 次
- Explanations are a Means to an End: Decision Theoretic Explanation EvaluationZiyang Guo, Berk Ustun, Jessica HullmanICML 2026
它引用的顶会 Paper21
- Does the Whole Exceed its Parts? The Effect of AI Explanations on Complementary Team PerformanceGagan Bansal, Tongshuang Wu, Joyce Zhou, Raymond Fok 等CHI 2021 · 被引用 713 次
- Manipulating and Measuring Model InterpretabilityForough Poursabzi-Sangdeh, Daniel G. Goldstein, Jake M. Hofman, Jennifer Wortman Vaughan 等CHI 2021 · 被引用 663 次
- Evaluating Explainable AI: Which Algorithmic Explanations Help Users Predict Model Behavior?Peter Hase, Mohit BansalACL 2020 · 被引用 216 次
- Is the Most Accurate AI the Best Teammate? Optimizing AI for TeamworkGagan Bansal, Besmira Nushi, Ece Kamar, Eric Horvitz 等AAAI 2021 · 被引用 185 次
- Counterfactual Explanations Can Be ManipulatedDylan Slack, Anna Hilgard, Himabindu Lakkaraju, Sameer SinghNeurIPS 2021 · 被引用 182 次
相关 Paper
- Decoding AI's Nudge: A Unified Framework to Predict Human Behavior in AI-Assisted Decision MakingZhuoyan Li, Zhuoran Lu, Ming YinAAAI 2024 · 被引用 23 次
- Uncalibrated Models Can Improve Human-AI CollaborationKailas Vodrahalli, Tobias Gerstenberg, James Y. ZouNeurIPS 2022 · 被引用 47 次
- Will You Accept the AI Recommendation? Predicting Human Behavior in AI-Assisted Decision MakingXinru Wang, Zhuoran Lu, Ming YinWWW 2022 · 被引用 58 次
- Modeling Human Trust and Reliance in AI-Assisted Decision Making: A Markovian ApproachZhuoyan Li, Zhuoran Lu, Ming YinAAAI 2023 · 被引用 28 次
- Watch Out for Updates: Understanding the Effects of Model Explanation Updates in AI-Assisted Decision MakingXinru Wang, Ming YinCHI 2023 · 被引用 34 次
