Actively Testing Your Model While It Learns: Realizing Label-Efficient Learning in Practice
Dayou Yu, Weishi Shi, Qi Yu
Abstract
In active learning (AL), we focus on reducing the data annotation cost from the model training perspective. However, “testing”, which often refers to the model evaluation process of using empirical risk to estimate the intractable true generalization risk, also requires data annotations. The annotation cost for “testing” (model evaluation) is under-explored. Even in works that study active model evaluation or active testing (AT), the learning and testing ends are disconnected. In this paper, we propose a novel active testing while learning (ATL) framework that integrates active learning with active testing. ATL provides an unbiased sample-efficient estimation of the model risk during active learning. It leverages test samples annotated from different periods of a dynamic active learning process to achieve fair model evaluations based on a theoretically guaranteed optimal integration of different test samples. Periodic testing also enables effective early-stopping to further save the total annotation cost. ATL further integrates an “active feedback” mechanism, which is inspired by human learning, where the teacher (active tester) provides immediate guidance given by the prior performance of the student (active learner). Our theoretical result reveals that active feedback maintains the label complexity of the integrated learning-testing objective, while improving the model’s generalization capability. We study the realistic setting where we maximize the performance gain from choosing “testing” samples for feedback without sacrificing the risk estimation accuracy. An agnostic-style analysis and empirical evaluations on real-world datasets demonstrate that the ATL framework can effectively improve the annotation efficiency of both active learning and evaluation tasks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers1
Ask how each one uses itBuilds on6
- Generalization bounds for deep convolutional neural networksPhilip M. Long, Hanie SedghiICLR 2020 · 102 citations
- On Statistical Bias In Active Learning: How and When to Fix ItSebastian Farquhar, Yarin Gal, Tom RainforthICLR 2021 · 96 citations
- Active Testing: Sample-Efficient Model EvaluationJannik Kossen, Sebastian Farquhar, Yarin Gal, Tom RainforthICML 2021 · 81 citations
- Active Surrogate Estimators: An Active Learning Approach to Label-Efficient Model EvaluationJannik Kossen, Sebastian Farquhar, Yarin Gal, Thomas RainforthNeurIPS 2022 · 36 citations
- Adaptive Region-Based Active LearningCorinna Cortes, Giulia DeSalvo, Claudio Gentile, Mehryar Mohri et al.ICML 2020 · 24 citations
Related papers
- Instance-wise Supervision-level Optimization in Active LearningShinnosuke Matsuo, Riku Togashi, Ryoma Bise, Seiichi Uchida et al.CVPR 2025
- Towards Cost-Effective Learning: A Synergy of Semi-Supervised and Active LearningTianxiang Yin, Ningzhong Liu, Han SunCVPR 2025
- Tracing Training Progress: Dynamic Influence Based Selection for Active LearningTianjiao Wan, Kele Xu, Long Lan, Zijian Gao et al.ACM MM 2024 · 3 citations
- Effortless Active Labeling for Long-Term Test-Time AdaptationGuowei Wang, Changxing DingCVPR 2025
- Not All Out-of-Distribution Data Are Harmful to Open-Set Active LearningYang Yang, Yuxuan Zhang, Xin Song, Yi XuNeurIPS 2023 · 48 citations
