Active Learning with Neural Networks: Insights from Nonparametric Statistics
Yinglun Zhu, Robert Nowak
Abstract
Deep neural networks have great representation power, but typically require large numbers of training examples. This motivates deep active learning methods that can significantly reduce the amount of labeled training data. Empirical successes of deep active learning have been recently reported in the literature, however, rigorous label complexity guarantees of deep active learning have remained elusive. This constitutes a significant gap between theory and practice. This paper tackles this gap by providing the first near-optimal label complexity guarantees for deep active learning. The key insight is to study deep active learning from the nonparametric classification perspective. Under standard low noise conditions, we show that active learning with neural networks can provably achieve the minimax label complexity, up to disagreement coefficient and other logarithmic terms. When equipped with an abstention option, we further develop an efficient deep active learning algorithm that achieves label complexity, without any low noise assumptions. We also provide extensions of our results beyond the commonly studied Sobolev/Hölder spaces and develop label complexity guarantees for learning in Radon spaces, which have recently been proposed as natural function spaces associated with neural networks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 642cc6c6-be95-4242-9380-750a3df01fc6Cited by top-tier papers6
- Neural Active Learning Beyond BanditsYikun Ban, Ishika Agarwal, Ziwei Wu, Yada Zhu et al.ICLR 2024 · 14 citations
- Test-Time Matching: Unlocking Compositional Reasoning in Multimodal ModelsYinglun Zhu, Jiancheng Zhang, Fuzhi TangICLR 2026 · 5 citations
- Budgeted Active Experimentation for Treatment Effect Estimation from Observational and Randomized DataJiacan Gao, Xinyan Su, Mingyuan Ma, Yiyan HUANG et al.ICML 2026 · 1 citation
- Learning to Help in Multi-Class SettingsYu Wu, Yansong Li, Zeyu Dong, Nitya Sathyavageeswaran et al.ICLR 2025
- Active Learning with Foundation Model Priors: Efficient Learning under Class ImbalanceJiancheng Zhang, Meiqing Li, Qi Zhang, Yinglun ZhuICML 2026
Builds on7
- Deep Batch Active Learning by Diverse, Uncertain Gradient Lower BoundsJordan T. Ash, Chicheng Zhang, Akshay Krishnamurthy, John Langford et al.ICLR 2020 · 974 citations
- Batch Active Learning at ScaleGui Citovsky, Giulia DeSalvo, Claudio Gentile, Lazaros Karydas et al.NeurIPS 2021 · 220 citations
- A Function Space View of Bounded Norm Infinite Width ReLU Nets: The Multivariate CaseGreg Ongie, Rebecca Willett, Daniel Soudry, Nathan SrebroICLR 2020 · 172 citations
- SIMILAR: Submodular Information Measures Based Active Learning In Realistic ScenariosSuraj Kothawade, Nathan Beck, KrishnaTeja Killamsetty, Rishabh K. IyerNeurIPS 2021 · 138 citations
- Gone Fishing: Neural Active Learning with Fisher EmbeddingsJordan T. Ash, Surbhi Goel, Akshay Krishnamurthy, Sham M. KakadeNeurIPS 2021 · 124 citations
Related papers
- Efficient Active Learning with AbstentionYinglun Zhu, Robert NowakNeurIPS 2022 · 27 citations
- Deep Active Learning by Leveraging Training DynamicsHaonan Wang, Wei Huang, Ziwei Wu, Hanghang Tong et al.NeurIPS 2022 · 49 citations
- Neural Active Learning with Performance GuaranteesZhilei Wang, Pranjal Awasthi, Christoph Dann, Ayush Sekhari et al.NeurIPS 2021 · 26 citations
- Robust Regression of General ReLUs with QueriesIlias Diakonikolas, Daniel Kane, Mingchen MaNeurIPS 2025 · 1 citation
- The Human-AI Substitution game: active learning from a strategic labelerTom Yan, Chicheng ZhangICLR 2024
