Efficient Active Learning with Abstention
Yinglun Zhu, Robert Nowak
Abstract
The goal of active learning is to achieve the same accuracy achievable by passive learning, while using much fewer labels. Exponential savings in terms of label complexity have been proved in very special cases, but fundamental lower bounds show that such improvements are impossible in general. This suggests a need to explore alternative goals for active learning. Learning with abstention is one such alternative. In this setting, the active learning algorithm may abstain from prediction and incur an error that is marginally smaller than random guessing. We develop the first computationally efficient active learning algorithm with abstention. Our algorithm provably achieves label complexity, without any low noise conditions. Such performance guarantee reduces the label complexity by an exponential factor, relative to passive learning and active learning that is not allowed to abstain. Furthermore, our algorithm is guaranteed to only abstain on hard examples (where the true label distribution is close to a fair coin), a novel property we term proper abstention that also leads to a host of other desirable characteristics (e.g., recovering minimax guarantees in the standard setting, and avoiding the undesirable"noise-seeking"behavior often seen in active learning). We also provide novel extensions of our algorithm that achieve constant label complexity and deal with model misspecification.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers16
- Making RL with Preference-based Feedback Efficient via RandomizationRunzhe Wu, Wen SunICLR 2024 · 44 citations
- Contextual Bandits and Imitation Learning with Preference-Based Active QueriesAyush Sekhari, Karthik Sridharan, Wen Sun, Runzhe WuNeurIPS 2023 · 18 citations
- Promises and Pitfalls of Threshold-based Auto-labelingHarit Vishwakarma, Heguang Lin, Frederic Sala, Ramya Korlakai VinayakNeurIPS 2023 · 16 citations
- Active Learning with Neural Networks: Insights from Nonparametric StatisticsYinglun Zhu, Robert NowakNeurIPS 2022 · 15 citations
- Selective Sampling and Imitation Learning via Online RegressionAyush Sekhari, Karthik Sridharan, Wen Sun, Runzhe WuNeurIPS 2023 · 15 citations
Builds on3
- Beyond UCB: Optimal and Efficient Contextual Bandits with Regression OraclesDylan J. Foster, Alexander RakhlinICML 2020 · 241 citations
- Learning with Good Feature Representations in Bandits and in RL with a Generative ModelTor Lattimore, Csaba Szepesvári, Gellért WeiszICML 2020 · 181 citations
- Improved Algorithms for Agnostic Pool-based Active ClassificationJulian Katz-Samuels, Jifan Zhang, Lalit Jain, Kevin JamiesonICML 2021 · 26 citations
Related papers
- Learning with Labeling Induced AbstentionsKareem Amin, Giulia DeSalvo, Afshin RostamizadehNeurIPS 2021 · 9 citations
- The Human-AI Substitution game: active learning from a strategic labelerTom Yan, Chicheng ZhangICLR 2024
- Metric-Fair Active LearningJie Shen, Nan Cui, Jing WangICML 2022 · 11 citations
- Towards optimally abstaining from prediction with OOD test examplesAdam Kalai, Varun KanadeNeurIPS 2021 · 1 citation
- SEL-BALD: Deep Bayesian Active Learning with Selective LabelsRuijiang Gao, Mingzhang Yin, Maytal Saar-TsechanskyNeurIPS 2024 · 4 citations
