Uncalibrated Models Can Improve Human-AI Collaboration
Kailas Vodrahalli, Tobias Gerstenberg, James Y. Zou
摘要
In many practical applications of AI, an AI model is used as a decision aid for human users. The AI provides advice that a human (sometimes) incorporates into their decision-making process. The AI advice is often presented with some measure of"confidence"that the human can use to calibrate how much they depend on or trust the advice. In this paper, we present an initial exploration that suggests showing AI models as more confident than they actually are, even when the original AI is well-calibrated, can improve human-AI performance (measured as the accuracy and confidence of the human's final prediction after seeing the AI advice). We first train a model to predict human incorporation of AI advice using data from thousands of human-AI interactions. This enables us to explicitly estimate how to transform the AI's prediction confidence, making the AI uncalibrated, in order to improve the final human prediction. We empirically validate our results across four different tasks--dealing with images, text and tabular data--involving hundreds of human participants. We further support our findings with simulation analysis. Our findings suggest the importance of jointly optimizing the human-AI system as opposed to the standard paradigm of optimizing the AI model alone.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper18
- Large Language Models Must Be Taught to Know What They Don't KnowSanyam Kapoor, Nate Gruver, Manley Roberts, Katie Collins 等NeurIPS 2024 · 被引用 124 次
- Improving Expert Predictions with Conformal PredictionEleni Straitouri, Lequn Wang, Nastaran Okati, Manuel Gomez RodriguezICML 2023 · 被引用 56 次
- Human-Aligned Calibration for AI-Assisted Decision MakingNina Corvelo Benz, Manuel Gomez RodriguezNeurIPS 2023 · 被引用 45 次
- Designing for Appropriate Reliance: The Roles of AI Uncertainty Presentation, Initial User Decision, and User Demographics in AI-Assisted Decision-MakingShiye Cao, Anqi Liu, Chien-Ming HuangCSCW 2024 · 被引用 38 次
- To Rely or Not to Rely? Evaluating Interventions for Appropriate Reliance on Large Language ModelsJessica Y. Bo, Sophia Wan, Ashton AndersonCHI 2025 · 被引用 31 次
它引用的顶会 Paper4
- Consistent Estimators for Learning to Defer to an ExpertHussein Mozannar, David A. SontagICML 2020 · 被引用 267 次
- Is the Most Accurate AI the Best Teammate? Optimizing AI for TeamworkGagan Bansal, Besmira Nushi, Ece Kamar, Eric Horvitz 等AAAI 2021 · 被引用 185 次
- CheXplain: Enabling Physicians to Explore and Understand Data-Driven, AI-Enabled Medical Imaging AnalysisYao Xie, Melody Chen, David Kao, Ge Gao 等CHI 2020 · 被引用 137 次
- Human Reliance on Machine Learning Models When Performance Feedback is Limited: Heuristics and RisksZhuoran Lu, Ming YinCHI 2021 · 被引用 123 次
相关 Paper
- Too Sure for Our Own Good: A User Study on AI Confidence and Human RelianceCaterina Fregosi, Lucia Vicente, Andrea Campagner, Federico CabitzaAAAI 2026 · 被引用 1 次
- As Confidence Aligns: Understanding the Effect of AI Confidence on Human Self-confidence in Human-AI Decision MakingJingshu Li, Yitian Yang, Q. Vera Liao, Junti Zhang 等CHI 2025 · 被引用 56 次
- Does Explainable Artificial Intelligence Improve Human Decision-Making?Yasmeen Alufaisan, Laura R. Marusich, Jonathan Z. Bakdash, Yan Zhou 等AAAI 2021 · 被引用 135 次
- Improving Human-AI Collaboration With Descriptions of AI BehaviorÁngel Alexander Cabrera, Adam Perer, Jason I. HongCSCW 2023 · 被引用 85 次
- "Are You Really Sure?" Understanding the Effects of Human Self-Confidence Calibration in AI-Assisted Decision MakingShuai Ma, Xinru Wang, Ying Lei, Chuhan Shi 等CHI 2024 · 被引用 54 次
