Uncertainty Herding: One Active Learning Method for All Label Budgets
Wonho Bae, Danica J. Sutherland, Gabriel L. Oliveira
Abstract
Most active learning research has focused on methods which perform well when many labels are available, but can be dramatically worse than random selection when label budgets are small. Other methods have focused on the low-budget regime, but do poorly as label budgets increase. As the line between "low" and "high" budgets varies by problem, this is a serious issue in practice. We propose uncertainty coverage, an objective which generalizes a variety of low- and high-budget objectives, as well as natural, hyperparameter-light methods to smoothly interpolate between low- and high-budget regimes. We call greedy optimization of the estimate Uncertainty Herding; this simple method is computationally fast, and we prove that it nearly optimizes the distribution-level coverage. In experimental validation across a variety of active learning tasks, our proposal matches or beats state-of-the-art performance in essentially all cases; it is the only method of which we are aware that reliably works well in both low- and high-budget settings.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 14ac683d-4fcc-4d50-90b6-0f1c63a468d3Cited by top-tier papers2
- Cleaning the Pool: Progressive Filtering of Unlabeled Pools in Deep Active LearningDenis Huseljic, Marek Herde, Lukas Rauch, Paul Hahn et al.CVPR 2026 · 2 citations
- Diffusion-Driven Two-Stage Active Learning for Low-Budget Semantic SegmentationJeongin Kim, Wonho Bae, YouLee Han, Giyeong Oh et al.NeurIPS 2025
Builds on14
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Training data-efficient image transformers & distillation through attentionHugo Touvron, Matthieu Cord, Matthijs Douze, Francisco Massa et al.ICML 2021 · 8,974 citations
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou et al.ICCV 2021 · 8,921 citations
- Moment Matching for Multi-Source Domain AdaptationXingchao Peng, Qinxun Bai, Xide Xia, Zijun Huang et al.ICCV 2019 · 2,239 citations
- Deep Batch Active Learning by Diverse, Uncertain Gradient Lower BoundsJordan T. Ash, Chicheng Zhang, Akshay Krishnamurthy, John Langford et al.ICLR 2020 · 974 citations
Related papers
- Convergence of Uncertainty Sampling for Active LearningAnant Raj, Francis R. BachICML 2022 · 41 citations
- Active Learning Through a Covering LensOfer Yehuda, Avihu Dekel, Guy Hacohen, Daphna WeinshallNeurIPS 2022 · 102 citations
- How to Select Which Active Learning Strategy is Best Suited for Your Specific Problem and BudgetGuy Hacohen, Daphna WeinshallNeurIPS 2023 · 23 citations
- Robust Sampling for Active Statistical InferencePuheng Li, Tijana Zrnic, Emmanuel J. CandèsNeurIPS 2025 · 6 citations
- CoverICL: Selective Annotation for In-Context Learning via Active Graph CoverageCostas Mavromatis, Balasubramaniam Srinivasan, Zhengyuan Shen, Jiani Zhang et al.EMNLP 2024
