Augmenting Clinical Decision-Making with an Interactive and Interpretable AI Copilot: A Real-World User Study with Clinicians in Nephrology and Obstetrics
Yinghao Zhu, Dehao Sui, Zixiang Wang, Xuning Hu, Lei Gu, Yifan Qi, Tianchen Wu, Ling Wang, Yuan Wei, Wen Tang, Zhihan Cui, Yasha Wang
Abstract
Clinician skepticism toward opaque AI hinders adoption in high-stakes healthcare. We present AICare, an interactive and interpretable AI copilot for collaborative clinical decision-making. By analyzing longitudinal electronic health records, AICare grounds dynamic risk predictions in scrutable visualizations and LLM-driven diagnostic recommendations. Through a within-subjects counterbalanced study with 16 clinicians across nephrology and obstetrics, we comprehensively evaluated AICare using objective measures (task completion time and error rate), subjective assessments (NASA-TLX, SUS, and confidence ratings), and semi-structured interviews. Our findings indicate AICare’s reduced cognitive workload. Beyond performance metrics, qualitative analysis reveals that trust is actively constructed through verification, with interaction strategies diverging by expertise: junior clinicians used the system as cognitive scaffolding to structure their analysis, while experts engaged in adversarial verification to challenge the AI’s logic. This work offers design implications for creating AI systems that function as transparent partners, accommodating diverse reasoning styles to augment rather than replace clinical judgment.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e12bfa38-c403-432f-95cc-08fb45d20ab1Builds on15
- To Trust or to Think: Cognitive Forcing Functions Can Reduce Overreliance on AI in AI-assisted Decision-makingZana Buçinca, Maja Barbara Malaya, Krzysztof Z. GajosCSCW 2021 · 962 citations
- Re-examining Whether, Why, and How Human-AI Interaction Is Uniquely Difficult to DesignQian Yang, Aaron Steinfeld, Carolyn P. Rosé, John ZimmermanCHI 2020 · 604 citations
- A Human-Centered Evaluation of a Deep Learning System Deployed in Clinics for the Detection of Diabetic RetinopathyEmma Beede, Elizabeth Elliott Baylor, Fred Hersch, Anna Iurchenko et al.CHI 2020 · 589 citations
- "Brilliant AI Doctor" in Rural Clinics: Challenges in AI-Powered Clinical Decision Support System DeploymentDakuo Wang, Liuping Wang, Zhan Zhang, Ding Wang et al.CHI 2021 · 207 citations
- ConCare: Personalized Clinical Feature Embedding via Capturing the Healthcare ContextLiantao Ma, Chaohe Zhang, Yasha Wang, Wenjie Ruan et al.AAAI 2020 · 190 citations
Related papers
- Exploring the Future of AI in Clinical Collaboration: A Study on Tumor Board Case PreparationJiachen Li, Amanda K. Hall, Ruican Rachel Zhong, Selin S. Everett et al.CHI 2026 · 1 citation
- ColaCare: Enhancing Electronic Health Record Modeling through Large Language Model-Driven Multi-Agent CollaborationZixiang Wang, Yinghao Zhu, Huiya Zhao, Xiaochen Zheng et al.WWW 2025 · 34 citations
- GLEAN: Guideline-Grounded Evidence Accumulation for High-Stakes Agent VerificationYichi Zhang, Nabeel Seedat, Yinpeng Dong, Peng Cui et al.ICML 2026 · 3 citations
- Ignore, Trust, or Negotiate: Understanding Clinician Acceptance of AI-Based Treatment Recommendations in Health CareVenkatesh Sivaraman, Leigh A. Bukowski, Joel Levin, Jeremy M. Kahn et al.CHI 2023 · 126 citations
- CliCARE: Grounding Large Language Models in Clinical Guidelines for Decision Support over Longitudinal Cancer Electronic Health RecordsDongchen Li, Jitao Liang, Wei Li, Xiaoyu Wang et al.AAAI 2026 · 1 citation
