Multitask Learning with No Regret: from Improved Confidence Bounds to Active Learning
Pier Giuseppe Sessa, Pierre Laforgue, Nicolò Cesa-Bianchi, Andreas Krause
Abstract
Multitask learning is a powerful framework that enables one to simultaneously learn multiple related tasks by sharing information between them. Quantifying uncertainty in the estimated tasks is of pivotal importance for many downstream applications, such as online or active learning. In this work, we provide novel multitask confidence intervals in the challenging agnostic setting, i.e., when neither the similarity between tasks nor the tasks' features are available to the learner. The obtained intervals do not require i.i.d. data and can be directly applied to bound the regret in online learning. Through a refined analysis of the multitask information gain, we obtain new regret guarantees that, depending on a task similarity parameter, can significantly improve over treating tasks independently. We further propose a novel online learning algorithm that achieves such improved regret without knowing this parameter in advance, i.e., automatically adapting to task similarity. As a second key application of our results, we introduce a novel multitask active learning setup where several tasks must be simultaneously optimized, but only one of them can be queried for feedback by the learner at each round. For this problem, we design a no-regret algorithm that uses our confidence intervals to decide which task should be queried. Finally, we empirically validate our bounds and algorithms on synthetic and real-world (drug discovery) data. * equal contribution Preprint. Under review.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers2
- Multi-task Linear Regression without Eigenvalue Lower Bounds: Adaptivity, Robustness, and SafetySeok-Jin KimICML 2026
- Combinatorial Bandit Bayesian Optimization for Tensor OutputsJingru Huang, Haijie Xu, Jie Guo, Manrui Jiang et al.ICLR 2026
Builds on2
Related papers
- Thompson Sampling for Robust Transfer in Multi-Task BanditsZhi Wang, Chicheng Zhang, Kamalika ChaudhuriICML 2022 · 7 citations
- Distributed Primal-Dual Optimization for Online Multi-Task LearningPeng Yang, Ping LiAAAI 2020 · 6 citations
- Byzantine Resilient Distributed Multi-Task LearningJiani Li, Waseem Abbas, Xenofon D. KoutsoukosNeurIPS 2020 · 12 citations
- Efficient and Effective Multi-task Grouping via Meta Learning on Task CombinationsXiaozhuang Song, Shun Zheng, Wei Cao, James J. Q. Yu et al.NeurIPS 2022 · 50 citations
- Stochastic Online Conformal Prediction with Semi-Bandit FeedbackHaosen Ge, Hamsa Bastani, Osbert BastaniICML 2025
