Improved Active Multi-Task Representation Learning via Lasso
Yiping Wang, Yifang Chen, Kevin Jamieson, Simon Shaolei Du
Abstract
To leverage the copious amount of data from source tasks and overcome the scarcity of the target task samples, representation learning based on multi-task pretraining has become a standard approach in many applications. However, up until now, most existing works design a source task selection strategy from a purely empirical perspective. Recently, gave the first active multi-task representation learning (A-MTRL) algorithm which adaptively samples from source tasks and can provably reduce the total sample complexity using the L2-regularized-target-source-relevance parameter . But their work is theoretically suboptimal in terms of total source sample complexity and is less practical in some real-world scenarios where sparse training source task selection is desired. In this paper, we address both issues. Specifically, we show the strict dominance of the L1-regularized-relevance-based (-based) strategy by giving a lower bound for the -based strategy. When is unknown, we propose a practical algorithm that uses the LASSO program to estimate . Our algorithm successfully recovers the optimal result in the known case. In addition to our sample complexity results, we also characterize the potential of our -based strategy in sample-cost-sensitive settings. Finally, we provide experiments on real-world computer vision datasets to illustrate the effectiveness of our proposed method.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0e5b0448-30c4-4615-9ebb-8aaf13637e88Cited by top-tier papers5
- Learning to Collaborate with Unknown Agents in the Absence of RewardZuyuan Zhang, Hanhan Zhou, Mahdi Imani, Taeyoung Lee et al.AAAI 2025 · 11 citations
- Interpretable Meta-Learning of Physical SystemsMatthieu Blanke, Marc LelargeICLR 2024 · 10 citations
- Sparse-to-dense Multimodal Image Registration via Multi-Task LearningKaining Zhang, Jiayi MaICML 2024 · 6 citations
- Guarantees for Nonlinear Representation Learning: Non-identical Covariates, Dependent Data, Fewer SamplesThomas T. C. K. Zhang, Bruce D. Lee, Ingvar M. Ziemann, George J. Pappas et al.ICML 2024 · 2 citations
- FloE: On-the-Fly MoE Inference on Memory-constrained GPUYuxin Zhou, Zheng Li, Jun Zhang, Jue Wang et al.ICML 2025
Builds on15
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Flamingo: a Visual Language Model for Few-Shot LearningJean-Baptiste Alayrac, Jeff Donahue, Pauline Luc, Antoine Miech et al.NeurIPS 2022 · 6,707 citations
- Exploiting Shared Representations for Personalized Federated LearningLiam Collins, Hamed Hassani, Aryan Mokhtari, Sanjay ShakkottaiICML 2021 · 1,081 citations
- Fine-Tuning can Distort Pretrained Features and Underperform Out-of-DistributionAnanya Kumar, Aditi Raghunathan, Robbie Matthew Jones, Tengyu Ma et al.ICLR 2022 · 911 citations
- Which Tasks Should Be Learned Together in Multi-task Learning?Trevor Standley, Amir Zamir, Dawn Chen, Leonidas J. Guibas et al.ICML 2020 · 651 citations
Related papers
- Active Multi-Task Representation LearningYifang Chen, Kevin Jamieson, Simon S. DuICML 2022 · 18 citations
- Active representation learning for general task space with applications in roboticsYifang Chen, Yingbing Huang, Simon S. Du, Kevin Jamieson et al.NeurIPS 2023 · 3 citations
- Adaptive Adversarial Multi-task Representation LearningYuren Mao, Weiwei Liu, Xuemin LinICML 2020 · 15 citations
- Provable Meta-Learning of Linear RepresentationsNilesh Tripuraneni, Chi Jin, Michael I. JordanICML 2021 · 218 citations
- Multi-Task Representation Alignment on Language Understanding: A Mutual Information PerspectiveDou Hu, Lingwei Wei, Hongjiang Xiao, Songlin Hu et al.ACL 2026
