Exploiting Human-AI Dependence for Learning to Defer
Zixi Wei, Yuzhou Cao, Lei Feng
摘要
The learning to defer (L2D) framework allows models to defer their decisions to human experts. For L2D, the Bayes optimality is the basic requirement of theoretical guarantees for the design of consistent surrogate loss functions, which requires the minimizer (i.e., learned classifier) by the surrogate loss to be the Bayes optimality. However, we find that the original form of Bayes optimality fails to consider the dependence between the model and the expert, and such a dependence could be further exploited to design a better consistent loss for L2D. In this paper, we provide a new formulation for the Bayes optimality called dependent Bayes optimality, which reveals the dependence pattern in determining whether to defer. Based on the dependent Bayes optimality, we further present a deferral principle for L2D. Following the guidance of the deferral principle, we propose a novel consistent surrogate loss. Comprehensive experimental results on both synthetic and real-world datasets demonstrate the superiority of our proposed method.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Deferring Concept Bottleneck Models: Learning to Defer Interventions to Inaccurate ExpertsAndrea Pugnana, Riccardo Massidda, Francesco Giannini, Pietro Barbiero 等NeurIPS 2025 · 被引用 11 次
- Human-AI Collaborative Uncertainty QuantificationSima Noorani, Shayan Kiyani, George Pappas, Hamed HassaniICML 2026 · 被引用 8 次
- Why Ask One When You Can Ask k? Learning-to-Defer to the Top-k ExpertsYannis Montreuil, Axel Carlier, Lai Xing Ng, Wei Tsang OoiICLR 2026 · 被引用 7 次
- Human-LLM Collaborative Feature Engineering for Tabular DataZhuoyan Li, Aditya Bansal, Jinzhao Li, Shishuang He 等ICLR 2026 · 被引用 2 次
- When More Experts Hurt: Underfitting in Multi-Expert Learning to DeferShuqi Liu, Yuzhou Cao, Lei Feng, Bo An 等ICML 2026 · 被引用 1 次
它引用的顶会 Paper13
- Learning with Noisy Labels Revisited: A Study Using Real-World Human AnnotationsJiaheng Wei, Zhaowei Zhu, Hao Cheng, Tongliang Liu 等ICLR 2022 · 被引用 338 次
- Provably Consistent Partial-Label LearningLei Feng, Jiaqi Lv, Bo Han, Miao Xu 等NeurIPS 2020 · 被引用 188 次
- Rethinking Calibration of Deep Neural Networks: Do Not Be Afraid of OverconfidenceDeng-Bao Wang, Lei Feng, Min-Ling ZhangNeurIPS 2021 · 被引用 177 次
- Two-Stage Learning to Defer with Multiple ExpertsAnqi Mao, Christopher Mohri, Mehryar Mohri, Yutao ZhongNeurIPS 2023 · 被引用 98 次
- Combining Human Predictions with Model Probabilities via Confusion Matrices and CalibrationGavin Kerrigan, Padhraic Smyth, Mark SteyversNeurIPS 2021 · 被引用 79 次
相关 Paper
- Mastering Multiple-Expert Routing: Realizable H-Consistency and Strong Guarantees for Learning to DeferAnqi Mao, Mehryar Mohri, Yutao ZhongICML 2025
- A Two-Stage Learning-to-Defer Approach for Multi-Task LearningYannis Montreuil, Yeo Shu Heng, Axel Carlier, Lai Xing Ng 等ICML 2025
- Sample Efficient Learning of Predictors that Complement HumansMohammad-Amin Charusaie, Hussein Mozannar, David A. Sontag, Samira SamadiICML 2022 · 被引用 52 次
- In Defense of Softmax Parametrization for Calibrated and Consistent Learning to DeferYuzhou Cao, Hussein Mozannar, Lei Feng, Hongxin Wei 等NeurIPS 2023 · 被引用 36 次
- Regression with Multi-Expert DeferralAnqi Mao, Mehryar Mohri, Yutao ZhongICML 2024 · 被引用 31 次
