Navigating the Pitfalls of Active Learning Evaluation: A Systematic Framework for Meaningful Performance Assessment
Carsten T. Lüth, Till J. Bungert, Lukas Klein, Paul F. Jaeger
摘要
Active Learning (AL) aims to reduce the labeling burden by interactively selecting the most informative samples from a pool of unlabeled data. While there has been extensive research on improving AL query methods in recent years, some studies have questioned the effectiveness of AL compared to emerging paradigms such as semi-supervised (Semi-SL) and self-supervised learning (Self-SL), or a simple optimization of classifier configurations. Thus, today's AL literature presents an inconsistent and contradictory landscape, leaving practitioners uncertain about whether and how to use AL in their tasks. In this work, we make the case that this inconsistency arises from a lack of systematic and realistic evaluation of AL methods. Specifically, we identify five key pitfalls in the current literature that reflect the delicate considerations required for AL evaluation. Further, we present an evaluation framework that overcomes these pitfalls and thus enables meaningful statements about the performance of AL methods. To demonstrate the relevance of our protocol, we present a large-scale empirical study and benchmark for image classification spanning various data sets, query methods, AL settings, and training paradigms. Our findings clarify the inconsistent picture in the literature and enable us to give hands-on recommendations for practitioners. The benchmark is hosted at https://github.com/IML-DKFZ/realistic-al .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- ValUES: A Framework for Systematic Validation of Uncertainty Estimation in Semantic SegmentationKim-Celine Kahl, Carsten T. Lüth, Maximilian Zenk, Klaus H. Maier-Hein 等ICLR 2024 · 被引用 28 次
- Addressing Pitfalls in the Evaluation of Uncertainty Estimation Methods for Natural Language GenerationMykyta Ielanskyi, Kajetan Schweighofer, Lukas Aichberger, Sepp HochreiterICLR 2026 · 被引用 10 次
- On the Fragility of Active Learners for Text ClassificationAbhishek Ghose, Emma NguyenEMNLP 2024 · 被引用 2 次
- Cleaning the Pool: Progressive Filtering of Unlabeled Pools in Deep Active LearningDenis Huseljic, Marek Herde, Lukas Rauch, Paul Hahn 等CVPR 2026 · 被引用 2 次
它引用的顶会 Paper15
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- FixMatch: Simplifying Semi-Supervised Learning with Consistency and ConfidenceKihyuk Sohn, David Berthelot, Nicholas Carlini, Zizhao Zhang 等NeurIPS 2020 · 被引用 5,129 次
- RandAugment: Practical Automated Data Augmentation with a Reduced Search SpaceEkin Dogus Cubuk, Barret Zoph, Jonathon Shlens, Quoc LeNeurIPS 2020 · 被引用 4,453 次
- Deep Batch Active Learning by Diverse, Uncertain Gradient Lower BoundsJordan T. Ash, Chicheng Zhang, Akshay Krishnamurthy, John Langford 等ICLR 2020 · 被引用 974 次
相关 Paper
- A Neural Pre-Conditioning Active Learning Algorithm to Reduce Label ComplexitySeo Taek Kong, Soomin Jeon, Dongbin Na, Jaewon Lee 等NeurIPS 2022 · 被引用 7 次
- To Label or Not to Label: PALM - a Predictive Model for Evaluating Sample Efficiency in Active Learning ModelsJulia Machnio, Mads Nielsen, Mostafa Mehdipour-GhaziICCV 2025 · 被引用 1 次
- Systematic comparison of semi-supervised and self-supervised learning for medical image classificationZhe Huang, Ruijie Jiang, Shuchin Aeron, Michael C. HughesCVPR 2024
- Querying Easily Flip-flopped Samples for Deep Active LearningSeong Jin Cho, Gwangsu Kim, Junghyun Lee, Jinwoo Shin 等ICLR 2024 · 被引用 8 次
- SIMILAR: Submodular Information Measures Based Active Learning In Realistic ScenariosSuraj Kothawade, Nathan Beck, KrishnaTeja Killamsetty, Rishabh K. IyerNeurIPS 2021 · 被引用 138 次
