Scalarization for Multi-Task and Multi-Domain Learning at Scale
Amelie Royer, Tijmen Blankevoort, Babak Ehteshami Bejnordi
Abstract
Training a single model on multiple input domains and/or output tasks allows for compressing information from multiple sources into a unified backbone hence improves model efficiency. It also enables potential positive knowledge transfer across tasks/domains, leading to improved accuracy and data-efficient training. However, optimizing such networks is a challenge, in particular due to discrepancies between the different tasks or domains: Despite several hypotheses and solutions proposed over the years, recent work has shown that uniform scalarization training, i.e., simply minimizing the average of the task losses, yields on-par performance with more costly SotA optimization methods. This raises the issue of how well we understand the training dynamics of multi-task and multi-domain networks. In this work, we first devise a large-scale unified analysis of multi-domain and multi-task learning to better understand the dynamics of scalarization across varied task/domain combinations and model sizes. Following these insights, we then propose to leverage population-based training to efficiently search for the optimal scalarization weights when dealing with a large number of tasks or domains.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e56ef953-1231-489b-9872-4baa84edfdaeCited by top-tier papers12
- Smooth Tchebycheff Scalarization for Multi-Objective OptimizationXi Lin, Xiaoyuan Zhang, Zhiyuan Yang, Fei Liu et al.ICML 2024 · 48 citations
- AMAGO-2: Breaking the Multi-Task Barrier in Meta-Reinforcement Learning with TransformersJake Grigsby, Justin Sasek, Samyak Parajuli, Daniel Adebi et al.NeurIPS 2024 · 19 citations
- DeltaFlow: An Efficient Multi-frame Scene Flow Estimation MethodQingwen Zhang, Xiaomeng Zhu, Yushan Zhang, Yixi Cai et al.NeurIPS 2025 · 8 citations
- An Information-theoretic Multi-task Representation Learning Framework for Natural Language UnderstandingDou Hu, Lingwei Wei, Wei Zhou, Songlin HuAAAI 2025 · 3 citations
- Scalable Multitask Learning Using Gradient-based Estimation of Task AffinityDongyue Li, Aneesh Sharma, Hongyang R. ZhangKDD 2024 · 3 citations
Builds on18
- RandAugment: Practical Automated Data Augmentation with a Reduced Search SpaceEkin Dogus Cubuk, Barret Zoph, Jonathon Shlens, Quoc LeNeurIPS 2020 · 4,453 citations
- Gradient Surgery for Multi-Task LearningTianhe Yu, Saurabh Kumar, Abhishek Gupta, Sergey Levine et al.NeurIPS 2020 · 2,261 citations
- Moment Matching for Multi-Source Domain AdaptationXingchao Peng, Qinxun Bai, Xide Xia, Zijun Huang et al.ICCV 2019 · 2,239 citations
- Conflict-Averse Gradient Descent for Multi-task learningBo Liu, Xingchao Liu, Xiaojie Jin, Peter Stone et al.NeurIPS 2021 · 686 citations
- Which Tasks Should Be Learned Together in Multi-task Learning?Trevor Standley, Amir Zamir, Dawn Chen, Leonidas J. Guibas et al.ICML 2020 · 651 citations
Related papers
- In Defense of the Unitary Scalarization for Deep Multi-Task LearningVitaly Kurin, Alessandro De Palma, Ilya Kostrikov, Shimon Whiteson et al.NeurIPS 2022 · 96 citations
- MDL-NAS: A Joint Multi-domain Learning Framework for Vision TransformerShiguang Wang, Tao Xie, Jian Cheng, Xingcheng Zhang et al.CVPR 2023
- Cross-Domain Collaborative Normalization via Structural KnowledgeHaifeng Xia, Zhengming DingAAAI 2022 · 5 citations
- A High-Dimensional Statistical Method for Optimizing Transfer Quantities in Multi-Source Transfer LearningQingyue Zhang, Haohao Fu, Guanbo Huang, Yaoyuan Liang et al.NeurIPS 2025 · 5 citations
- Deep Elastic Networks With Model Selection for Multi-Task LearningChanho Ahn, Eunwoo Kim, Songhwai OhICCV 2019 · 56 citations
