Exploiting Human-AI Dependence for Learning to Defer
Zixi Wei, Yuzhou Cao, Lei Feng
Abstract
The learning to defer (L2D) framework allows models to defer their decisions to human experts. For L2D, the Bayes optimality is the basic requirement of theoretical guarantees for the design of consistent surrogate loss functions, which requires the minimizer (i.e., learned classifier) by the surrogate loss to be the Bayes optimality. However, we find that the original form of Bayes optimality fails to consider the dependence between the model and the expert, and such a dependence could be further exploited to design a better consistent loss for L2D. In this paper, we provide a new formulation for the Bayes optimality called dependent Bayes optimality, which reveals the dependence pattern in determining whether to defer. Based on the dependent Bayes optimality, we further present a deferral principle for L2D. Following the guidance of the deferral principle, we propose a novel consistent surrogate loss. Comprehensive experimental results on both synthetic and real-world datasets demonstrate the superiority of our proposed method.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b457bd4d-cd65-467f-bd5d-7a1e6d9005dbCited by top-tier papers8
- Deferring Concept Bottleneck Models: Learning to Defer Interventions to Inaccurate ExpertsAndrea Pugnana, Riccardo Massidda, Francesco Giannini, Pietro Barbiero et al.NeurIPS 2025 · 11 citations
- Human-AI Collaborative Uncertainty QuantificationSima Noorani, Shayan Kiyani, George Pappas, Hamed HassaniICML 2026 · 8 citations
- Why Ask One When You Can Ask k? Learning-to-Defer to the Top-k ExpertsYannis Montreuil, Axel Carlier, Lai Xing Ng, Wei Tsang OoiICLR 2026 · 7 citations
- Human-LLM Collaborative Feature Engineering for Tabular DataZhuoyan Li, Aditya Bansal, Jinzhao Li, Shishuang He et al.ICLR 2026 · 2 citations
- When More Experts Hurt: Underfitting in Multi-Expert Learning to DeferShuqi Liu, Yuzhou Cao, Lei Feng, Bo An et al.ICML 2026 · 1 citation
Builds on13
- Learning with Noisy Labels Revisited: A Study Using Real-World Human AnnotationsJiaheng Wei, Zhaowei Zhu, Hao Cheng, Tongliang Liu et al.ICLR 2022 · 338 citations
- Provably Consistent Partial-Label LearningLei Feng, Jiaqi Lv, Bo Han, Miao Xu et al.NeurIPS 2020 · 188 citations
- Rethinking Calibration of Deep Neural Networks: Do Not Be Afraid of OverconfidenceDeng-Bao Wang, Lei Feng, Min-Ling ZhangNeurIPS 2021 · 177 citations
- Two-Stage Learning to Defer with Multiple ExpertsAnqi Mao, Christopher Mohri, Mehryar Mohri, Yutao ZhongNeurIPS 2023 · 98 citations
- Combining Human Predictions with Model Probabilities via Confusion Matrices and CalibrationGavin Kerrigan, Padhraic Smyth, Mark SteyversNeurIPS 2021 · 79 citations
Related papers
- Mastering Multiple-Expert Routing: Realizable H-Consistency and Strong Guarantees for Learning to DeferAnqi Mao, Mehryar Mohri, Yutao ZhongICML 2025
- A Two-Stage Learning-to-Defer Approach for Multi-Task LearningYannis Montreuil, Yeo Shu Heng, Axel Carlier, Lai Xing Ng et al.ICML 2025
- Sample Efficient Learning of Predictors that Complement HumansMohammad-Amin Charusaie, Hussein Mozannar, David A. Sontag, Samira SamadiICML 2022 · 52 citations
- In Defense of Softmax Parametrization for Calibrated and Consistent Learning to DeferYuzhou Cao, Hussein Mozannar, Lei Feng, Hongxin Wei et al.NeurIPS 2023 · 36 citations
- Regression with Multi-Expert DeferralAnqi Mao, Mehryar Mohri, Yutao ZhongICML 2024 · 31 citations
