Efficient Active Learning with Abstention
Yinglun Zhu, Robert Nowak
摘要
The goal of active learning is to achieve the same accuracy achievable by passive learning, while using much fewer labels. Exponential savings in terms of label complexity have been proved in very special cases, but fundamental lower bounds show that such improvements are impossible in general. This suggests a need to explore alternative goals for active learning. Learning with abstention is one such alternative. In this setting, the active learning algorithm may abstain from prediction and incur an error that is marginally smaller than random guessing. We develop the first computationally efficient active learning algorithm with abstention. Our algorithm provably achieves label complexity, without any low noise conditions. Such performance guarantee reduces the label complexity by an exponential factor, relative to passive learning and active learning that is not allowed to abstain. Furthermore, our algorithm is guaranteed to only abstain on hard examples (where the true label distribution is close to a fair coin), a novel property we term proper abstention that also leads to a host of other desirable characteristics (e.g., recovering minimax guarantees in the standard setting, and avoiding the undesirable"noise-seeking"behavior often seen in active learning). We also provide novel extensions of our algorithm that achieve constant label complexity and deal with model misspecification.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper16
- Making RL with Preference-based Feedback Efficient via RandomizationRunzhe Wu, Wen SunICLR 2024 · 被引用 44 次
- Contextual Bandits and Imitation Learning with Preference-Based Active QueriesAyush Sekhari, Karthik Sridharan, Wen Sun, Runzhe WuNeurIPS 2023 · 被引用 18 次
- Promises and Pitfalls of Threshold-based Auto-labelingHarit Vishwakarma, Heguang Lin, Frederic Sala, Ramya Korlakai VinayakNeurIPS 2023 · 被引用 16 次
- Active Learning with Neural Networks: Insights from Nonparametric StatisticsYinglun Zhu, Robert NowakNeurIPS 2022 · 被引用 15 次
- Selective Sampling and Imitation Learning via Online RegressionAyush Sekhari, Karthik Sridharan, Wen Sun, Runzhe WuNeurIPS 2023 · 被引用 15 次
它引用的顶会 Paper3
- Beyond UCB: Optimal and Efficient Contextual Bandits with Regression OraclesDylan J. Foster, Alexander RakhlinICML 2020 · 被引用 241 次
- Learning with Good Feature Representations in Bandits and in RL with a Generative ModelTor Lattimore, Csaba Szepesvári, Gellért WeiszICML 2020 · 被引用 181 次
- Improved Algorithms for Agnostic Pool-based Active ClassificationJulian Katz-Samuels, Jifan Zhang, Lalit Jain, Kevin JamiesonICML 2021 · 被引用 26 次
相关 Paper
- Learning with Labeling Induced AbstentionsKareem Amin, Giulia DeSalvo, Afshin RostamizadehNeurIPS 2021 · 被引用 9 次
- The Human-AI Substitution game: active learning from a strategic labelerTom Yan, Chicheng ZhangICLR 2024
- Metric-Fair Active LearningJie Shen, Nan Cui, Jing WangICML 2022 · 被引用 11 次
- Towards optimally abstaining from prediction with OOD test examplesAdam Kalai, Varun KanadeNeurIPS 2021 · 被引用 1 次
- SEL-BALD: Deep Bayesian Active Learning with Selective LabelsRuijiang Gao, Mingzhang Yin, Maytal Saar-TsechanskyNeurIPS 2024 · 被引用 4 次
