Diversity Enhanced Active Learning with Strictly Proper Scoring Rules
Wei Tan, Lan Du, Wray L. Buntine
摘要
We study acquisition functions for active learning (AL) for text classification. The Expected Loss Reduction (ELR) method focuses on a Bayesian estimate of the reduction in classification error, recently updated with Mean Objective Cost of Uncertainty (MOCU). We convert the ELR framework to estimate the increase in (strictly proper) scores like log probability or negative mean square error, which we call Bayesian Estimate of Mean Proper Scores (BEMPS 2 ). We also prove convergence results borrowing techniques used with MOCU. In order to allow better experimentation with the new acquisition functions, we develop a complementary batch AL algorithm, which encourages diversity in the vector of expected changes in scores for unlabelled data. To allow high performance text classifiers, we combine ensembling and dynamic validation set construction on pretrained language models. Extensive experimental evaluation then explores how these different acquisition functions perform. The results show that the use of mean square error and log probability with BEMPS yields robust acquisition functions, which consistently outperform the others tested.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- No Change, No Gain: Empowering Graph Neural Networks with Expected Model Change Maximization for Active LearningZixing Song, Yifei Zhang, Irwin KingNeurIPS 2023 · 被引用 21 次
- Improved Algorithms for Neural Active LearningYikun Ban, Yuheng Zhang, Hanghang Tong, Arindam Banerjee 等NeurIPS 2022 · 被引用 18 次
- Neural Active Learning Beyond BanditsYikun Ban, Ishika Agarwal, Ziwei Wu, Yada Zhu 等ICLR 2024 · 被引用 14 次
- AUC Maximization for Low-Resource Named Entity RecognitionNgoc Dang Nguyen, Wei Tan, Lan Du, Wray L. Buntine 等AAAI 2023 · 被引用 12 次
- Counterfactual Active Learning for Out-of-Distribution GeneralizationXun Deng, Wenjie Wang, Fuli Feng, Hanwang Zhang 等ACL 2023 · 被引用 10 次
它引用的顶会 Paper8
- Deep Batch Active Learning by Diverse, Uncertain Gradient Lower BoundsJordan T. Ash, Chicheng Zhang, Akshay Krishnamurthy, John Langford 等ICLR 2020 · 被引用 974 次
- Bayesian Deep Learning and a Probabilistic Perspective of GeneralizationAndrew Gordon Wilson, Pavel IzmailovNeurIPS 2020 · 被引用 845 次
- Variational Adversarial Active LearningSamarth Sinha, Sayna Ebrahimi, Trevor DarrellICCV 2019 · 被引用 662 次
- Pitfalls of In-Domain Uncertainty Estimation and Ensembling in Deep LearningArsenii Ashukha, Alexander Lyzhov, Dmitry Molchanov, Dmitry P. VetrovICLR 2020 · 被引用 354 次
- Cold-start Active Learning through Self-supervised Language ModelingMichelle Yuan, Hsuan-Tien Lin, Jordan L. Boyd-GraberEMNLP 2020 · 被引用 128 次
相关 Paper
- Uncertainty-aware Active Learning for Optimal Bayesian ClassifierGuang Zhao, Edward R. Dougherty, Byung-Jun Yoon, Francis J. Alexander 等ICLR 2021 · 被引用 43 次
- Harnessing the Power of Beta Scoring in Deep Active Learning for Multi-Label Text ClassificationWei Tan, Ngoc Dang Nguyen, Lan Du, Wray L. BuntineAAAI 2024 · 被引用 5 次
- Active Learning in Bayesian Neural Networks with Balanced Entropy Learning PrincipleJae Oh WooICLR 2023 · 被引用 2 次
- An Information-Theoretic Framework for Unifying Active Learning ProblemsQuoc Phong Nguyen, Bryan Kian Hsiang Low, Patrick JailletAAAI 2021 · 被引用 23 次
- Deep Bayesian Active Learning for Preference Modeling in Large Language ModelsLuckeciano Carvalho Melo, Panagiotis Tigas, Alessandro Abate, Yarin GalNeurIPS 2024 · 被引用 25 次
