Improving Human-AI Collaboration With Descriptions of AI Behavior
Ángel Alexander Cabrera, Adam Perer, Jason I. Hong
Abstract
People work with AI systems to improve their decision making, but often under- or over-rely on AI predictions and perform worse than they would have unassisted. To help people appropriately rely on AI aids, we propose showing them behavior descriptions, details of how AI systems perform on subgroups of instances. We tested the efficacy of behavior descriptions through user studies with 225 participants in three distinct domains: fake review detection, satellite image classification, and bird classification. We found that behavior descriptions can increase human-AI accuracy through two mechanisms: helping people identify AI failures and increasing people's reliance on the AI when it is more accurate. These findings highlight the importance of people's mental models in human-AI collaboration and show that informing people of high-level AI behaviors can significantly improve AI-assisted decision making.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2b6a3c7b-ad36-47ce-a09c-dae56c8c1526Cited by top-tier papers18
- The Impact of Imperfect XAI on Human-AI Decision-MakingKatelyn Morrison, Philipp Spitzer, Violet Turri, Michelle Feng et al.CSCW 2024 · 60 citations
- "Are You Really Sure?" Understanding the Effects of Human Self-Confidence Calibration in AI-Assisted Decision MakingShuai Ma, Xinru Wang, Ying Lei, Chuhan Shi et al.CHI 2024 · 54 citations
- Effect of Explanation Conceptualisations on Trust in AI-assisted Credibility AssessmentSaumya Pareek, Niels van Berkel, Eduardo Velloso, Jorge GonçalvesCSCW 2024 · 39 citations
- Contrastive Explanations That Anticipate Human Misconceptions Can Improve Human Decision-Making SkillsZana Buçinca, Siddharth Swaroop, Amanda E. Paluch, Finale Doshi-Velez et al.CHI 2025 · 31 citations
- Perceptions of Sentient AI and Other Digital Minds: Evidence from the AI, Morality, and Sentience (AIMS) SurveyJacy Reese Anthis, Janet V. T. Pauketat, Ali Ladak, Aikaterina ManoliCHI 2025 · 24 citations
Builds on14
- To Trust or to Think: Cognitive Forcing Functions Can Reduce Overreliance on AI in AI-assisted Decision-makingZana Buçinca, Maja Barbara Malaya, Krzysztof Z. GajosCSCW 2021 · 962 citations
- Does the Whole Exceed its Parts? The Effect of AI Explanations on Complementary Team PerformanceGagan Bansal, Tongshuang Wu, Joyce Zhou, Raymond Fok et al.CHI 2021 · 713 citations
- Manipulating and Measuring Model InterpretabilityForough Poursabzi-Sangdeh, Daniel G. Goldstein, Jake M. Hofman, Jennifer Wortman Vaughan et al.CHI 2021 · 663 citations
- Interpreting Interpretability: Understanding Data Scientists' Use of Interpretability Tools for Machine LearningHarmanpreet Kaur, Harsha Nori, Samuel Jenkins, Rich Caruana et al.CHI 2020 · 541 citations
- "An Ideal Human": Expectations of AI Teammates in Human-AI TeamingRui Zhang, Nathan J. McNeese, Guo Freeman, Geoff MusickCSCW 2020 · 222 citations
Related papers
- Uncalibrated Models Can Improve Human-AI CollaborationKailas Vodrahalli, Tobias Gerstenberg, James Y. ZouNeurIPS 2022 · 47 citations
- Does More Advice Help? The Effects of Second Opinions in AI-Assisted Decision MakingZhuoran Lu, Dakuo Wang, Ming YinCSCW 2024 · 36 citations
- Does Explainable Artificial Intelligence Improve Human Decision-Making?Yasmeen Alufaisan, Laura R. Marusich, Jonathan Z. Bakdash, Yan Zhou et al.AAAI 2021 · 135 citations
- Human Reliance on Machine Learning Models When Performance Feedback is Limited: Heuristics and RisksZhuoran Lu, Ming YinCHI 2021 · 123 citations
- "Why is 'Chicago' deceptive?" Towards Building Model-Driven Tutorials for HumansVivian Lai, Han Liu, Chenhao TanCHI 2020 · 113 citations
