'The AI is uncertain, so am I. What now?': Navigating Shortcomings of Uncertainty Representations in Human-AI Collaboration with Capability-focused Guidance
Ulrike Schäfer, Lars Sipos, Claudia Müller-Birn
摘要
As AI becomes increasingly relevant, especially in high-stakes domains such as healthcare, it is important to investigate which approaches can improve human-AI collaboration and, if so, why. Current research focuses primarily on technically available approaches, such as explainable AI (XAI), often overlooking human needs. This study bridges this gap by adopting a well-established technical approach - model uncertainty representations - by considering users' familiarity with the format and numeracy skills. Despite being provided with uncertainty representations, users may still struggle to handle uncertain decisions. Thus, we introduce an educational approach that communicates the capabilities of humans and the AI system to users, supplementing the uncertainty representations. We conducted a pre-registered, between-subjects user study to determine whether these approaches resulted in improved human-AI team performance, mediated by the user's mental model of the AI. Our findings indicate that solely providing uncertainty representations does not improve team performance or the user's mental model in comparison to only providing AI recommendations. However, incorporating capability-focused guidance alongside uncertainty representations significantly enhances correct self-reliance and, to some extent, overall team performance. Our additional exploratory analyses suggest that factors such as task uncertainty, case difficulty, and case type, rather than numeracy skills, the need for cognition or familiarity, can influence team performance. We discuss these factors in detail, provide practical implications, and suggest directions for further research. This work contributes to the CSCW discourse by demonstrating how technical approaches can be augmented with educational approaches to enhance human-AI collaboration in decision-making tasks.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper3
- Designing Staged Evaluation Workflows for LLMs: Integrating Domain Experts, Lay Users, and Model-Generated Evaluation CriteriaAnnalisa Szymanski, Simret Araya Gebreegziabher, Oghenemaro Anuyah, Ronald A. Metoyer 等CHI 2026 · 被引用 1 次
- From Use to Oversight: How Mental Models Influence User Behavior and Output in AI Writing AssistantsShalaleh Rismani, Su Lin Blodgett, Q. Vera Liao, Alexandra Olteanu 等CHI 2026 · 被引用 1 次
- People Can Accurately Predict Behavior of Complex Algorithms That Are Available, Compact, and Aligned CSCW031Lindsay Popowski, Helena Vasconcelos, Ignacio Javier Fernandez, Chijioke Chinaza Mgbahurike 等CSCW 2026
相关 Paper
- Designing for Appropriate Reliance: The Roles of AI Uncertainty Presentation, Initial User Decision, and User Demographics in AI-Assisted Decision-MakingShiye Cao, Anqi Liu, Chien-Ming HuangCSCW 2024 · 被引用 38 次
- The Impact of Imperfect XAI on Human-AI Decision-MakingKatelyn Morrison, Philipp Spitzer, Violet Turri, Michelle Feng 等CSCW 2024 · 被引用 60 次
- Does the Whole Exceed its Parts? The Effect of AI Explanations on Complementary Team PerformanceGagan Bansal, Tongshuang Wu, Joyce Zhou, Raymond Fok 等CHI 2021 · 被引用 713 次
- Teaching Humans When to Defer to a Classifier via ExemplarsHussein Mozannar, Arvind Satyanarayan, David A. SontagAAAI 2022 · 被引用 49 次
- The Utility of Explainable AI in Ad Hoc Human-Machine TeamingRohan R. Paleja, Muyleng Ghuy, Nadun Ranawaka Arachchige, Reed Jensen 等NeurIPS 2021 · 被引用 103 次
