Differentiable Learning Under Triage
Nastaran Okati, Abir De, Manuel Gomez-Rodriguez
Abstract
Multiple lines of evidence suggest that predictive models may benefit from algorithmic triage. Under algorithmic triage, a predictive model does not predict all instances but instead defers some of them to human experts. However, the interplay between the prediction accuracy of the model and the human experts under algorithmic triage is not well understood. In this work, we start by formally characterizing under which circumstances a predictive model may benefit from algorithmic triage. In doing so, we also demonstrate that models trained for full automation may be suboptimal under triage. Then, given any model and desired level of triage, we show that the optimal triage policy is a deterministic threshold rule in which triage decisions are derived deterministically by thresholding the difference between the model and human errors on a per-instance level. Building upon these results, we introduce a practical gradient-based algorithm that is guaranteed to find a sequence of triage policies and predictive models of increasing performance. Experiments on a wide variety of supervised learning tasks using synthetic and real data from two important applications -- content moderation and scientific discovery -- illustrate our theoretical results and show that the models and triage policies provided by our gradient-based algorithm outperform those provided by several competitive baselines.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4c4a656e-ce96-4507-9372-91f9abe14fe5Cited by top-tier papers33
- Two-Stage Learning to Defer with Multiple ExpertsAnqi Mao, Christopher Mohri, Mehryar Mohri, Yutao ZhongNeurIPS 2023 · 98 citations
- Combining Human Predictions with Model Probabilities via Confusion Matrices and CalibrationGavin Kerrigan, Padhraic Smyth, Mark SteyversNeurIPS 2021 · 79 citations
- Calibrated Learning to Defer with One-vs-All ClassifiersRajeev Verma, Eric T. NalisnickICML 2022 · 76 citations
- Improving Expert Predictions with Conformal PredictionEleni Straitouri, Lequn Wang, Nastaran Okati, Manuel Gomez RodriguezICML 2023 · 56 citations
- Sample Efficient Learning of Predictors that Complement HumansMohammad-Amin Charusaie, Hussein Mozannar, David A. Sontag, Samira SamadiICML 2022 · 52 citations
Builds on3
- Consistent Estimators for Learning to Defer to an ExpertHussein Mozannar, David A. SontagICML 2020 · 267 citations
- Regression under Human AssistanceAbir De, Paramita Koley, Niloy Ganguly, Manuel Gomez-RodriguezAAAI 2020 · 73 citations
- Classification Under Human AssistanceAbir De, Nastaran Okati, Ali Zarezade, Manuel Gomez RodriguezAAAI 2021 · 60 citations
Related papers
- Learning to Defer with Limited Expert PredictionsPatrick Hemmer, Lukas Thede, Michael Vössing, Johannes Jakubik et al.AAAI 2023 · 28 citations
- Human-AI Collaboration via Conditional Delegation: A Case Study of Content ModerationVivian Lai, Samuel Carton, Rajat Bhatnagar, Q. Vera Liao et al.CHI 2022 · 135 citations
- Wait! There's a Way Out: A Decision Mechanism for Forecasting Conversational DerailmentLaerdon Kim, Vivian Nguyen, Cristian Danescu-Niculescu-MizilACL 2026
- Deferring Concept Bottleneck Models: Learning to Defer Interventions to Inaccurate ExpertsAndrea Pugnana, Riccardo Massidda, Francesco Giannini, Pietro Barbiero et al.NeurIPS 2025 · 11 citations
- Human Assisted Learning by Evolutionary Multi-Objective OptimizationDan-Xuan Liu, Xin Mu, Chao QianAAAI 2023 · 9 citations
