Auditing for Human Expertise
Rohan Alur, Loren Laine, Darrick K. Li, Manish Raghavan, Devavrat Shah, Dennis L. Shung
摘要
High-stakes prediction tasks (e.g., patient diagnosis) are often handled by trained human experts. A common source of concern about automation in these settings is that experts may exercise intuition that is difficult to model and/or have access to information (e.g., conversations with a patient) that is simply unavailable to a would-be algorithm. This raises a natural question whether human experts add value which could not be captured by an algorithmic predictor. We develop a statistical framework under which we can pose this question as a natural hypothesis test. Indeed, as our framework highlights, detecting human expertise is more subtle than simply comparing the accuracy of expert predictions to those made by a particular learning algorithm. Instead, we propose a simple procedure which tests whether expert predictions are statistically independent from the outcomes of interest after conditioning on the available inputs ('features'). A rejection of our test thus suggests that human experts may add value to any algorithm trained on the available data, and has direct implications for whether human-AI 'complementarity' is achievable in a given prediction task. We highlight the utility of our procedure using admissions data collected from the emergency department of a large academic hospital system, where we show that physicians' admit/discharge decisions for patients with acute gastrointestinal bleeding (AGIB) appear to be incorporating information that is not available to a standard algorithmic screening tool. This is despite the fact that the screening tool is arguably more accurate than physicians' discretionary decisions, highlighting that -even absent normative concerns about accountability or interpretabilityaccuracy is insufficient to justify algorithmic automation. Preprint. Under review.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Human Expertise in Algorithmic PredictionRohan Alur, Manish Raghavan, Devavrat ShahNeurIPS 2024 · 被引用 18 次
- When Are Two Lists Better than One?: Benefits and Harms in Joint Decision-MakingKate Donahue, Sreenivas Gollapudi, Kostas KolliasAAAI 2024 · 被引用 8 次
- Unveiling the Uncertainty in Embodied and Operational Carbon of Large AI Models through a Probabilistic Carbon Accounting ModelXiaoyang Zhang, Fang He, Yang Deng, Dan WangNeurIPS 2025 · 被引用 3 次
- Large Language Models in Peer-Run Community Behavioral Health Services: Understanding Peer Specialists and Service Users' Perspectives on Opportunities, Risks, and Mitigation StrategiesCindy Peng, Megan Chai, Gao Mo, Naveen Raman 等CHI 2026 · 被引用 2 次
- The Value of Information in Human-AI Decision-makingZiyang Guo, Yifan Wu, Jason D. Hartline, Jessica HullmanICLR 2026
它引用的顶会 Paper3
- Does the Whole Exceed its Parts? The Effect of AI Explanations on Complementary Team PerformanceGagan Bansal, Tongshuang Wu, Joyce Zhou, Raymond Fok 等CHI 2021 · 被引用 713 次
- Performative PredictionJuan C. Perdomo, Tijana Zrnic, Celestine Mendler-Dünner, Moritz HardtICML 2020 · 被引用 422 次
- Consistent Estimators for Learning to Defer to an ExpertHussein Mozannar, David A. SontagICML 2020 · 被引用 267 次
相关 Paper
- Toward Supporting Perceptual Complementarity in Human-AI Collaboration via Reflection on UnobservablesKenneth Holstein, Maria De-Arteaga, Lakshmi Tumati, Yanghuidi ChengCSCW 2023 · 被引用 37 次
- A Case for Humans-in-the-Loop: Decisions in the Presence of Erroneous Algorithmic ScoresMaria De-Arteaga, Riccardo Fogliato, Alexandra ChouldechovaCHI 2020 · 被引用 176 次
- Rethinking Human-AI Collaboration in Complex Medical Decision Making: A Case Study in Sepsis DiagnosisShao Zhang, Jianing Yu, Xuhai Xu, Changchang Yin 等CHI 2024 · 被引用 95 次
- Human-Algorithmic Interaction Using a Large Language Model-Augmented Artificial Intelligence Clinical Decision Support SystemNiroop Channa Rajashekar, Yeo Eun Shin, Yuan Pu, Sunny Chung 等CHI 2024 · 被引用 85 次
- Does Explainable Artificial Intelligence Improve Human Decision-Making?Yasmeen Alufaisan, Laura R. Marusich, Jonathan Z. Bakdash, Yan Zhou 等AAAI 2021 · 被引用 135 次
