Active Learning in Bayesian Neural Networks with Balanced Entropy Learning Principle
Jae Oh Woo
Abstract
Acquiring labeled data is challenging in many machine learning applications with limited budgets. Active learning gives a procedure to select the most informative data points and improve data efficiency by reducing the cost of labeling. The infomax learning principle maximizing mutual information such as BALD has been successful and widely adapted in various active learning applications. However, this pool-based specific objective inherently introduces a redundant selection and further requires a high computational cost for batch selection. In this paper, we design and propose a new uncertainty measure, Balanced Entropy Acquisition (BalEntAcq), which captures the information balance between the uncertainty of underlying softmax probability and the label variable. To do this, we approximate each marginal distribution by Beta distribution. Beta approximation enables us to formulate BalEntAcq as a ratio between an augmented entropy and the marginalized joint entropy. The closed-form expression of BalEntAcq facilitates parallelization by estimating two parameters in each marginal Beta distribution. BalEntAcq is a purely standalone measure without requiring any relational computations with other data points. Nevertheless, BalEntAcq captures a well-diversified selection near the decision boundary with a margin, unlike other existing uncertainty measures such as BALD, Entropy, or Mean Standard Deviation (MeanSD). Finally, we demonstrate that our balanced entropy learning principle with BalEntAcq 1 consistently outperforms well-known linearly scalable active learning methods, including a recently proposed PowerBALD, a simple but diversified version of BALD, by showing experimental results obtained from MNIST, CIFAR-100, SVHN, and TinyImageNet datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3cb7b1e3-f088-4459-8ee2-a0dd5cac6d01Cited by top-tier papers4
- Querying Easily Flip-flopped Samples for Deep Active LearningSeong Jin Cho, Gwangsu Kim, Junghyun Lee, Jinwoo Shin et al.ICLR 2024 · 8 citations
- Unsupervised Accuracy Estimation of Deep Visual Models using Domain-Adaptive Adversarial Perturbation without Source SamplesJoonHo Lee, Jae Oh Woo, Hankyu Moon, Kwonho LeeICCV 2023 · 5 citations
- Improving Instruction Following in Language Models through Proxy-Based Uncertainty EstimationJoonHo Lee, Jae Oh Woo, Juree Seok, Parisa Hassanzadeh et al.ICML 2024 · 4 citations
- Diffusion-Driven Two-Stage Active Learning for Low-Budget Semantic SegmentationJeongin Kim, Wonho Bae, YouLee Han, Giyeong Oh et al.NeurIPS 2025
Builds on20
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec et al.NeurIPS 2020 · 9,171 citations
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou et al.ICCV 2021 · 8,921 citations
- Unsupervised Learning of Visual Features by Contrasting Cluster AssignmentsMathilde Caron, Ishan Misra, Julien Mairal, Priya Goyal et al.NeurIPS 2020 · 5,249 citations
- Big Self-Supervised Models are Strong Semi-Supervised LearnersTing Chen, Simon Kornblith, Kevin Swersky, Mohammad Norouzi et al.NeurIPS 2020 · 2,611 citations
Related papers
- Scalable Batch-Mode Deep Bayesian Active Learning via Equivalence Class AnnealingRenyu Zhang, Aly A. Khan, Robert L. Grossman, Yuxin ChenICLR 2023 · 1 citation
- Beam Search Optimized Batch Bayesian Active LearningJingyu Sun, Hongjie Zhai, Osamu Saisho, Susumu TakeuchiAAAI 2023 · 2 citations
- Diversity Enhanced Active Learning with Strictly Proper Scoring RulesWei Tan, Lan Du, Wray L. BuntineNeurIPS 2021 · 40 citations
- Uncertainty for Active Learning on GraphsDominik Fuchsgruber, Tom Wollschläger, Bertrand Charpentier, Antonio Oroz et al.ICML 2024 · 17 citations
- Enhancing Deep Batch Active Learning for Regression with Imperfect Data Guided SelectionYinjie Min, Furong Xu, Xinyao Li, Changliang Zou et al.NeurIPS 2025 · 1 citation
