Lune

NeurIPS2025Top-tier venue

Near-Exponential Savings for Population Mean Estimation with Active Learning

Julian M. Morimoto, Jacob S. Goldin, Daniel E. Ho

2025Year

Abstract

We study the problem of efficiently estimating the mean of a k-class random variable, Y , using a limited number of labels, N , in settings where the analyst has access to auxiliary information (i.e.: covariates) X that may be informative about Y . We propose an active learning algorithm ("PartiBandits"

, where c > 0 is a constant and ν is the risk of the Bayes-optimal classifier. PartiBandits is essentially a two-stage algorithm. In the first stage, it learns a partition of the unlabeled data that shrinks the average conditional variance of Y . In the second stage it uses a UCB-style subroutine ("WarmStart-UCB") to request labels from each stratum round-by-round. Both the main algorithm's and the subroutine's convergence rates are minimax optimal in classical settings. PartiBandits bridges the UCB and disagreement-based approaches to active learning despite these two approaches being designed to tackle very different tasks. We illustrate our methods through simulation using nationwide electronic health records. Our methods can be implemented using the PartiBandits package in R.

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

Builds on1

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines