Actively Testing Your Model While It Learns: Realizing Label-Efficient Learning in Practice
Dayou Yu, Weishi Shi, Qi Yu
摘要
In active learning (AL), we focus on reducing the data annotation cost from the model training perspective. However, “testing”, which often refers to the model evaluation process of using empirical risk to estimate the intractable true generalization risk, also requires data annotations. The annotation cost for “testing” (model evaluation) is under-explored. Even in works that study active model evaluation or active testing (AT), the learning and testing ends are disconnected. In this paper, we propose a novel active testing while learning (ATL) framework that integrates active learning with active testing. ATL provides an unbiased sample-efficient estimation of the model risk during active learning. It leverages test samples annotated from different periods of a dynamic active learning process to achieve fair model evaluations based on a theoretically guaranteed optimal integration of different test samples. Periodic testing also enables effective early-stopping to further save the total annotation cost. ATL further integrates an “active feedback” mechanism, which is inspired by human learning, where the teacher (active tester) provides immediate guidance given by the prior performance of the student (active learner). Our theoretical result reveals that active feedback maintains the label complexity of the integrated learning-testing objective, while improving the model’s generalization capability. We study the realistic setting where we maximize the performance gain from choosing “testing” samples for feedback without sacrificing the risk estimation accuracy. An agnostic-style analysis and empirical evaluations on real-world datasets demonstrate that the ATL framework can effectively improve the annotation efficiency of both active learning and evaluation tasks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper6
- Generalization bounds for deep convolutional neural networksPhilip M. Long, Hanie SedghiICLR 2020 · 被引用 102 次
- On Statistical Bias In Active Learning: How and When to Fix ItSebastian Farquhar, Yarin Gal, Tom RainforthICLR 2021 · 被引用 96 次
- Active Testing: Sample-Efficient Model EvaluationJannik Kossen, Sebastian Farquhar, Yarin Gal, Tom RainforthICML 2021 · 被引用 81 次
- Active Surrogate Estimators: An Active Learning Approach to Label-Efficient Model EvaluationJannik Kossen, Sebastian Farquhar, Yarin Gal, Thomas RainforthNeurIPS 2022 · 被引用 36 次
- Adaptive Region-Based Active LearningCorinna Cortes, Giulia DeSalvo, Claudio Gentile, Mehryar Mohri 等ICML 2020 · 被引用 24 次
相关 Paper
- Instance-wise Supervision-level Optimization in Active LearningShinnosuke Matsuo, Riku Togashi, Ryoma Bise, Seiichi Uchida 等CVPR 2025
- Towards Cost-Effective Learning: A Synergy of Semi-Supervised and Active LearningTianxiang Yin, Ningzhong Liu, Han SunCVPR 2025
- Tracing Training Progress: Dynamic Influence Based Selection for Active LearningTianjiao Wan, Kele Xu, Long Lan, Zijian Gao 等ACM MM 2024 · 被引用 3 次
- Effortless Active Labeling for Long-Term Test-Time AdaptationGuowei Wang, Changxing DingCVPR 2025
- Not All Out-of-Distribution Data Are Harmful to Open-Set Active LearningYang Yang, Yuxuan Zhang, Xin Song, Yi XuNeurIPS 2023 · 被引用 48 次
