Diversity Enhanced Active Learning with Strictly Proper Scoring Rules
Wei Tan, Lan Du, Wray L. Buntine
Abstract
We study acquisition functions for active learning (AL) for text classification. The Expected Loss Reduction (ELR) method focuses on a Bayesian estimate of the reduction in classification error, recently updated with Mean Objective Cost of Uncertainty (MOCU). We convert the ELR framework to estimate the increase in (strictly proper) scores like log probability or negative mean square error, which we call Bayesian Estimate of Mean Proper Scores (BEMPS 2 ). We also prove convergence results borrowing techniques used with MOCU. In order to allow better experimentation with the new acquisition functions, we develop a complementary batch AL algorithm, which encourages diversity in the vector of expected changes in scores for unlabelled data. To allow high performance text classifiers, we combine ensembling and dynamic validation set construction on pretrained language models. Extensive experimental evaluation then explores how these different acquisition functions perform. The results show that the use of mean square error and log probability with BEMPS yields robust acquisition functions, which consistently outperform the others tested.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d30b26b3-8178-425d-8cdf-1fb7fff29a5aCited by top-tier papers10
- No Change, No Gain: Empowering Graph Neural Networks with Expected Model Change Maximization for Active LearningZixing Song, Yifei Zhang, Irwin KingNeurIPS 2023 · 21 citations
- Improved Algorithms for Neural Active LearningYikun Ban, Yuheng Zhang, Hanghang Tong, Arindam Banerjee et al.NeurIPS 2022 · 18 citations
- Neural Active Learning Beyond BanditsYikun Ban, Ishika Agarwal, Ziwei Wu, Yada Zhu et al.ICLR 2024 · 14 citations
- AUC Maximization for Low-Resource Named Entity RecognitionNgoc Dang Nguyen, Wei Tan, Lan Du, Wray L. Buntine et al.AAAI 2023 · 12 citations
- Counterfactual Active Learning for Out-of-Distribution GeneralizationXun Deng, Wenjie Wang, Fuli Feng, Hanwang Zhang et al.ACL 2023 · 10 citations
Builds on8
- Deep Batch Active Learning by Diverse, Uncertain Gradient Lower BoundsJordan T. Ash, Chicheng Zhang, Akshay Krishnamurthy, John Langford et al.ICLR 2020 · 974 citations
- Bayesian Deep Learning and a Probabilistic Perspective of GeneralizationAndrew Gordon Wilson, Pavel IzmailovNeurIPS 2020 · 845 citations
- Variational Adversarial Active LearningSamarth Sinha, Sayna Ebrahimi, Trevor DarrellICCV 2019 · 662 citations
- Pitfalls of In-Domain Uncertainty Estimation and Ensembling in Deep LearningArsenii Ashukha, Alexander Lyzhov, Dmitry Molchanov, Dmitry P. VetrovICLR 2020 · 354 citations
- Cold-start Active Learning through Self-supervised Language ModelingMichelle Yuan, Hsuan-Tien Lin, Jordan L. Boyd-GraberEMNLP 2020 · 128 citations
Related papers
- Uncertainty-aware Active Learning for Optimal Bayesian ClassifierGuang Zhao, Edward R. Dougherty, Byung-Jun Yoon, Francis J. Alexander et al.ICLR 2021 · 43 citations
- Harnessing the Power of Beta Scoring in Deep Active Learning for Multi-Label Text ClassificationWei Tan, Ngoc Dang Nguyen, Lan Du, Wray L. BuntineAAAI 2024 · 5 citations
- Active Learning in Bayesian Neural Networks with Balanced Entropy Learning PrincipleJae Oh WooICLR 2023 · 2 citations
- An Information-Theoretic Framework for Unifying Active Learning ProblemsQuoc Phong Nguyen, Bryan Kian Hsiang Low, Patrick JailletAAAI 2021 · 23 citations
- Deep Bayesian Active Learning for Preference Modeling in Large Language ModelsLuckeciano Carvalho Melo, Panagiotis Tigas, Alessandro Abate, Yarin GalNeurIPS 2024 · 25 citations
