Achieving Minimax Rates in Pool-Based Batch Active Learning
Claudio Gentile, Zhilei Wang, Tong Zhang
摘要
We consider a batch active learning scenario where the learner adaptively issues batches of points to a labeling oracle. Sampling labels in batches is highly desirable in practice due to the smaller number of interactive rounds with the labeling oracle (often human beings). However, batch active learning typically pays the price of a reduced adaptivity, leading to suboptimal results. In this paper we propose a solution which requires a careful trade off between the informativeness of the queried points and their diversity. We theoretically investigate batch active learning in the practically relevant scenario where the unlabeled pool of data is available beforehand (pool-based active learning). We analyze a novel stage-wise greedy algorithm and show that, as a function of the label complexity, the excess risk of this algorithm matches the known minimax rates in standard statistical learning settings. Our results also exhibit a mild dependence on the batch size. These are the first theoretical results that employ careful trade offs between informativeness and diversity to rigorously quantify the statistical performance of batch active learning in the pool-based scenario.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Neural Active Learning Beyond BanditsYikun Ban, Ishika Agarwal, Ziwei Wu, Yada Zhu 等ICLR 2024 · 被引用 14 次
- Pessimistic Nonlinear Least-Squares Value Iteration for Offline Reinforcement LearningQiwei Di, Heyang Zhao, Jiafan He, Quanquan GuICLR 2024 · 被引用 9 次
- Actively Testing Your Model While It Learns: Realizing Label-Efficient Learning in PracticeDayou Yu, Weishi Shi, Qi YuNeurIPS 2023 · 被引用 5 次
- Active Learning based Structural InferenceAoran Wang, Jun PangICML 2023 · 被引用 2 次
- Adaptive Data Collection for Robust Learning Across Multiple DistributionsChengbo Zang, Mehmet Kerem Türkcan, Gil Zussman, Zoran Kostic 等ICML 2025
它引用的顶会 Paper9
- Deep Batch Active Learning by Diverse, Uncertain Gradient Lower BoundsJordan T. Ash, Chicheng Zhang, Akshay Krishnamurthy, John Langford 等ICLR 2020 · 被引用 974 次
- Batch Active Learning at ScaleGui Citovsky, Giulia DeSalvo, Claudio Gentile, Lazaros Karydas 等NeurIPS 2021 · 被引用 220 次
- SIMILAR: Submodular Information Measures Based Active Learning In Realistic ScenariosSuraj Kothawade, Nathan Beck, KrishnaTeja Killamsetty, Rishabh K. IyerNeurIPS 2021 · 被引用 138 次
- Improved Optimistic Algorithms for Logistic BanditsLouis Faury, Marc Abeille, Clément Calauzènes, Olivier FercoqICML 2020 · 被引用 127 次
- High-dimensional Experimental Design and Kernel BanditsRomain Camilleri, Kevin Jamieson, Julian Katz-SamuelsICML 2021 · 被引用 63 次
相关 Paper
- A Lagrangian Duality Approach to Active LearningJuan Elenter, Navid NaderiAlizadeh, Alejandro RibeiroNeurIPS 2022 · 被引用 31 次
- Active Classification with Few Queries under MisspecificationVasilis Kontonis, Mingchen Ma, Christos TzamosNeurIPS 2024 · 被引用 3 次
- Beam Search Optimized Batch Bayesian Active LearningJingyu Sun, Hongjie Zhai, Osamu Saisho, Susumu TakeuchiAAAI 2023 · 被引用 2 次
- Active Learning on Pre-trained Language Model with Task-Independent Triplet LossSeungmin Seo, Donghyun Kim, Youbin Ahn, Kyong-Ho LeeAAAI 2022 · 被引用 21 次
- Variational Adversarial Active LearningSamarth Sinha, Sayna Ebrahimi, Trevor DarrellICCV 2019 · 被引用 662 次
