Multitask Learning with No Regret: from Improved Confidence Bounds to Active Learning
Pier Giuseppe Sessa, Pierre Laforgue, Nicolò Cesa-Bianchi, Andreas Krause
摘要
Multitask learning is a powerful framework that enables one to simultaneously learn multiple related tasks by sharing information between them. Quantifying uncertainty in the estimated tasks is of pivotal importance for many downstream applications, such as online or active learning. In this work, we provide novel multitask confidence intervals in the challenging agnostic setting, i.e., when neither the similarity between tasks nor the tasks' features are available to the learner. The obtained intervals do not require i.i.d. data and can be directly applied to bound the regret in online learning. Through a refined analysis of the multitask information gain, we obtain new regret guarantees that, depending on a task similarity parameter, can significantly improve over treating tasks independently. We further propose a novel online learning algorithm that achieves such improved regret without knowing this parameter in advance, i.e., automatically adapting to task similarity. As a second key application of our results, we introduce a novel multitask active learning setup where several tasks must be simultaneously optimized, but only one of them can be queried for feedback by the learner at each round. For this problem, we design a no-regret algorithm that uses our confidence intervals to decide which task should be queried. Finally, we empirically validate our bounds and algorithms on synthetic and real-world (drug discovery) data. * equal contribution Preprint. Under review.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Multi-task Linear Regression without Eigenvalue Lower Bounds: Adaptivity, Robustness, and SafetySeok-Jin KimICML 2026
- Combinatorial Bandit Bayesian Optimization for Tensor OutputsJingru Huang, Haijie Xu, Jie Guo, Manrui Jiang 等ICLR 2026
它引用的顶会 Paper2
相关 Paper
- Thompson Sampling for Robust Transfer in Multi-Task BanditsZhi Wang, Chicheng Zhang, Kamalika ChaudhuriICML 2022 · 被引用 7 次
- Distributed Primal-Dual Optimization for Online Multi-Task LearningPeng Yang, Ping LiAAAI 2020 · 被引用 6 次
- Byzantine Resilient Distributed Multi-Task LearningJiani Li, Waseem Abbas, Xenofon D. KoutsoukosNeurIPS 2020 · 被引用 12 次
- Efficient and Effective Multi-task Grouping via Meta Learning on Task CombinationsXiaozhuang Song, Shun Zheng, Wei Cao, James J. Q. Yu 等NeurIPS 2022 · 被引用 50 次
- Stochastic Online Conformal Prediction with Semi-Bandit FeedbackHaosen Ge, Hamsa Bastani, Osbert BastaniICML 2025
