Why Ask One When You Can Ask k? Learning-to-Defer to the Top-k Experts
Yannis Montreuil, Axel Carlier, Lai Xing Ng, Wei Tsang Ooi
摘要
Existing Learning-to-Defer (L2D) frameworks are limited to single-expert deferral, forcing each query to rely on only one expert and preventing the use of collective expertise. We introduce the first framework for Top- Learning-to-Defer, which allocates queries to the most cost-effective entities. Our formulation unifies and strictly generalizes prior approaches, including the one-stage and two-stage regimes, selective prediction, and classical cascades. In particular, it recovers the usual Top-1 deferral rule as a special case while enabling principled collaboration with multiple experts when . We further propose Top- Learning-to-Defer, an adaptive variant that learns the optimal number of experts per query based on input difficulty, expert quality, and consultation cost. To enable practical learning, we develop a novel surrogate loss that is Bayes-consistent, -consistent in the one-stage setting, and -consistent in the two-stage setting. Crucially, this surrogate is independent of , allowing a single policy to be learned once and deployed flexibly across . Experiments across both regimes show that Top- and Top- deliver superior accuracy–cost trade-offs, opening a new direction for multi-expert deferral in L2D.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Linear-Core Surrogates: Smooth Loss Functions with Linear Rates for Classification and Structured PredictionMehryar Mohri, Yutao ZhongICML 2026 · 被引用 7 次
- Optimized Deferral for Imbalanced SettingsCorinna Cortes, Anqi Mao, Mehryar Mohri, Yutao ZhongICML 2026 · 被引用 7 次
- Mind the Gap: Structure-Aware Consistency in Preference LearningMehryar Mohri, Yutao ZhongICML 2026
它引用的顶会 Paper35
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Cross-Entropy Loss Functions: Theoretical Analysis and ApplicationsAnqi Mao, Mehryar Mohri, Yutao ZhongICML 2023 · 被引用 790 次
- Consistent Estimators for Learning to Defer to an ExpertHussein Mozannar, David A. SontagICML 2020 · 被引用 267 次
- Two-Stage Learning to Defer with Multiple ExpertsAnqi Mao, Christopher Mohri, Mehryar Mohri, Yutao ZhongNeurIPS 2023 · 被引用 98 次
- Combining Human Predictions with Model Probabilities via Confusion Matrices and CalibrationGavin Kerrigan, Padhraic Smyth, Mark SteyversNeurIPS 2021 · 被引用 79 次
相关 Paper
- Regression with Multi-Expert DeferralAnqi Mao, Mehryar Mohri, Yutao ZhongICML 2024 · 被引用 31 次
- Mastering Multiple-Expert Routing: Realizable H-Consistency and Strong Guarantees for Learning to DeferAnqi Mao, Mehryar Mohri, Yutao ZhongICML 2025
- Exploiting Human-AI Dependence for Learning to DeferZixi Wei, Yuzhou Cao, Lei FengICML 2024 · 被引用 15 次
- A Two-Stage Learning-to-Defer Approach for Multi-Task LearningYannis Montreuil, Yeo Shu Heng, Axel Carlier, Lai Xing Ng 等ICML 2025
- When More Experts Hurt: Underfitting in Multi-Expert Learning to DeferShuqi Liu, Yuzhou Cao, Lei Feng, Bo An 等ICML 2026 · 被引用 1 次
