Human Expertise in Algorithmic Prediction
Rohan Alur, Manish Raghavan, Devavrat Shah
Abstract
We introduce a novel framework for incorporating human expertise into algorithmic predictions. Our approach leverages human judgment to distinguish inputs which are algorithmically indistinguishable, or"look the same"to predictive algorithms. We argue that this framing clarifies the problem of human-AI collaboration in prediction tasks, as experts often form judgments by drawing on information which is not encoded in an algorithm's training data. Algorithmic indistinguishability yields a natural test for assessing whether experts incorporate this kind of"side information", and further provides a simple but principled method for selectively incorporating human feedback into algorithmic predictions. We show that this method provably improves the performance of any feasible algorithmic predictor and precisely quantify this improvement. We find empirically that although algorithms often outperform their human counterparts on average, human judgment can improve algorithmic predictions on specific instances (which can be identified ex-ante). In an X-ray classification task, we find that this subset constitutes nearly of the patient population. Our approach provides a natural way of uncovering this heterogeneity and thus enabling effective human-AI collaboration.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 190a973a-051e-4fd8-8725-2afb87d665afCited by top-tier papers6
- Tractable Agreement ProtocolsNatalie Collina, Surbhi Goel, Varun Gupta, Aaron RothSTOC 2025 · 10 citations
- Decision Theoretic Foundations for Experiments Evaluating Human DecisionsJessica Hullman, Alex Kale, Jason D. HartlineCHI 2025 · 6 citations
- How to Auto-optimize Prompts for Domain Tasks? Adaptive Prompting and Reasoning through Evolutionary Domain Knowledge AdaptationYang Zhao, Pu Wang, Hao (Frank) YangNeurIPS 2025 · 3 citations
- Human-LLM Collaborative Feature Engineering for Tabular DataZhuoyan Li, Aditya Bansal, Jinzhao Li, Shishuang He et al.ICLR 2026 · 2 citations
- AtC: Aggregate-then-Calibrate for Human-centered AssessmentZejun Xie, Xintong Li, Guang Wang, Desheng ZhangICLR 2026
Builds on22
- Does the Whole Exceed its Parts? The Effect of AI Explanations on Complementary Team PerformanceGagan Bansal, Tongshuang Wu, Joyce Zhou, Raymond Fok et al.CHI 2021 · 713 citations
- A Human-Centered Evaluation of a Deep Learning System Deployed in Clinics for the Detection of Diabetic RetinopathyEmma Beede, Elizabeth Elliott Baylor, Fred Hersch, Anna Iurchenko et al.CHI 2020 · 589 citations
- Performative PredictionJuan C. Perdomo, Tijana Zrnic, Celestine Mendler-Dünner, Moritz HardtICML 2020 · 422 citations
- Consistent Estimators for Learning to Defer to an ExpertHussein Mozannar, David A. SontagICML 2020 · 267 citations
- Is the Most Accurate AI the Best Teammate? Optimizing AI for TeamworkGagan Bansal, Besmira Nushi, Ece Kamar, Eric Horvitz et al.AAAI 2021 · 185 citations
Related papers
- Auditing for Human ExpertiseRohan Alur, Loren Laine, Darrick K. Li, Manish Raghavan et al.NeurIPS 2023 · 19 citations
- Learning to Defer with Limited Expert PredictionsPatrick Hemmer, Lukas Thede, Michael Vössing, Johannes Jakubik et al.AAAI 2023 · 28 citations
- Understanding Choice Independence and Error Types in Human-AI CollaborationAlexander Erlei, Abhinav Sharma, Ujwal GadirajuCHI 2024 · 25 citations
- A No Free Lunch Theorem for Human-AI CollaborationKenny Peng, Nikhil Garg, Jon M. KleinbergAAAI 2025 · 8 citations
- Human-AI Collaborative Bayesian OptimisationArun Kumar A. V., Santu Rana, Alistair Shilton, Svetha VenkateshNeurIPS 2022 · 24 citations
