Uncalibrated Models Can Improve Human-AI Collaboration
Kailas Vodrahalli, Tobias Gerstenberg, James Y. Zou
Abstract
In many practical applications of AI, an AI model is used as a decision aid for human users. The AI provides advice that a human (sometimes) incorporates into their decision-making process. The AI advice is often presented with some measure of"confidence"that the human can use to calibrate how much they depend on or trust the advice. In this paper, we present an initial exploration that suggests showing AI models as more confident than they actually are, even when the original AI is well-calibrated, can improve human-AI performance (measured as the accuracy and confidence of the human's final prediction after seeing the AI advice). We first train a model to predict human incorporation of AI advice using data from thousands of human-AI interactions. This enables us to explicitly estimate how to transform the AI's prediction confidence, making the AI uncalibrated, in order to improve the final human prediction. We empirically validate our results across four different tasks--dealing with images, text and tabular data--involving hundreds of human participants. We further support our findings with simulation analysis. Our findings suggest the importance of jointly optimizing the human-AI system as opposed to the standard paradigm of optimizing the AI model alone.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3059bf81-ae77-4d2e-8016-adb338c42df2Cited by top-tier papers18
- Large Language Models Must Be Taught to Know What They Don't KnowSanyam Kapoor, Nate Gruver, Manley Roberts, Katie Collins et al.NeurIPS 2024 · 124 citations
- Improving Expert Predictions with Conformal PredictionEleni Straitouri, Lequn Wang, Nastaran Okati, Manuel Gomez RodriguezICML 2023 · 56 citations
- Human-Aligned Calibration for AI-Assisted Decision MakingNina Corvelo Benz, Manuel Gomez RodriguezNeurIPS 2023 · 45 citations
- Designing for Appropriate Reliance: The Roles of AI Uncertainty Presentation, Initial User Decision, and User Demographics in AI-Assisted Decision-MakingShiye Cao, Anqi Liu, Chien-Ming HuangCSCW 2024 · 38 citations
- To Rely or Not to Rely? Evaluating Interventions for Appropriate Reliance on Large Language ModelsJessica Y. Bo, Sophia Wan, Ashton AndersonCHI 2025 · 31 citations
Builds on4
- Consistent Estimators for Learning to Defer to an ExpertHussein Mozannar, David A. SontagICML 2020 · 267 citations
- Is the Most Accurate AI the Best Teammate? Optimizing AI for TeamworkGagan Bansal, Besmira Nushi, Ece Kamar, Eric Horvitz et al.AAAI 2021 · 185 citations
- CheXplain: Enabling Physicians to Explore and Understand Data-Driven, AI-Enabled Medical Imaging AnalysisYao Xie, Melody Chen, David Kao, Ge Gao et al.CHI 2020 · 137 citations
- Human Reliance on Machine Learning Models When Performance Feedback is Limited: Heuristics and RisksZhuoran Lu, Ming YinCHI 2021 · 123 citations
Related papers
- Too Sure for Our Own Good: A User Study on AI Confidence and Human RelianceCaterina Fregosi, Lucia Vicente, Andrea Campagner, Federico CabitzaAAAI 2026 · 1 citation
- As Confidence Aligns: Understanding the Effect of AI Confidence on Human Self-confidence in Human-AI Decision MakingJingshu Li, Yitian Yang, Q. Vera Liao, Junti Zhang et al.CHI 2025 · 56 citations
- Does Explainable Artificial Intelligence Improve Human Decision-Making?Yasmeen Alufaisan, Laura R. Marusich, Jonathan Z. Bakdash, Yan Zhou et al.AAAI 2021 · 135 citations
- Improving Human-AI Collaboration With Descriptions of AI BehaviorÁngel Alexander Cabrera, Adam Perer, Jason I. HongCSCW 2023 · 85 citations
- "Are You Really Sure?" Understanding the Effects of Human Self-Confidence Calibration in AI-Assisted Decision MakingShuai Ma, Xinru Wang, Ying Lei, Chuhan Shi et al.CHI 2024 · 54 citations
