Mitigating Task Interference in Multi-Task Learning via Explicit Task Routing with Non-Learnable Primitives
Chuntao Ding, Zhichao Lu, Shangguang Wang, Ran Cheng, Vishnu Naresh Boddeti
摘要
Multi-task learning (MTL) seeks to learn a single model to accomplish multiple tasks by leveraging shared information among the tasks. Existing MTL models, however, have been known to suffer from negative interference among tasks. Efforts to mitigate task interference have focused on either loss/gradient balancing or implicit parameter partitioning with partial overlaps among the tasks. In this paper, we propose ETR-NLP to mitigate task interference through a synergistic combination of non-learnable primitives (NLPs) and explicit task routing (ETR). Our key idea is to employ non-learnable primitives to extract a diverse set of task-agnostic features and recombine them into a shared branch common to all tasks and explicit task-specific branches reserved for each task. The non-learnable primitives and the explicit decoupling of learnable parameters into shared and task-specific ones afford the flexibility needed for minimizing task interference. We evaluate the efficacy of ETR-NLP networks for both image-level classification and pixel-level dense prediction MTL problems. Experimental results indicate that ETR-NLP significantly outperforms state-of-the-art baselines with fewer learnable parameters and similar FLOPs across all datasets. Code is available at this URL.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Twin-Merging: Dynamic Integration of Modular Expertise in Model MergingZhenyi Lu, Chenghao Fan, Wei Wei, Xiaoye Qu 等NeurIPS 2024 · 被引用 139 次
- Towards Modular LLMs by Building and Reusing a Library of LoRAsOleksiy Ostapenko, Zhan Su, Edoardo M. Ponti, Laurent Charlin 等ICML 2024 · 被引用 70 次
- MoME: Mixture of Multimodal Experts for Generalist Multimodal Large Language ModelsLeyang Shen, Gongwei Chen, Rui Shao, Weili Guan 等NeurIPS 2024 · 被引用 55 次
- Denoising Task Routing for Diffusion ModelsByeongjun Park, Sangmin Woo, Hyojun Go, Jin-Young Kim 等ICLR 2024 · 被引用 26 次
- On Fairness of Task Arithmetic: The Role of Task VectorsLaura Gomezjurado Gonzalez, Hiroki Naganuma, Kotaro Yoshida, Takafumi Horie 等ICLR 2026 · 被引用 3 次
它引用的顶会 Paper8
- MetaFormer is Actually What You Need for VisionWeihao Yu, Mi Luo, Pan Zhou, Chenyang Si 等CVPR 2022 · 被引用 1,114 次
- Understanding and Improving Information Transfer in Multi-Task LearningSen Wu, Hongyang R. Zhang, Christopher RéICLR 2020 · 被引用 183 次
- Many Task Learning With Task RoutingGjorgji Strezoski, Nanne van Noord, Marcel WorringICCV 2019 · 被引用 112 次
- Stochastic Filter Groups for Multi-Task CNNs: Learning Specialist and Generalist Convolution KernelsFelix J. S. Bragman, Ryutaro Tanno, Sébastien Ourselin, Daniel C. Alexander 等ICCV 2019 · 被引用 97 次
- Maximum Roaming Multi-Task LearningLucas Pascal, Pietro Michiardi, Xavier Bost, Benoit Huet 等AAAI 2021 · 被引用 32 次
相关 Paper
- Improving Multi-Task Generalization via Regularizing Spurious CorrelationZiniu Hu, Zhe Zhao, Xinyang Yi, Tiansheng Yao 等NeurIPS 2022 · 被引用 46 次
- The Stem Cell Hypothesis: Dilemma behind Multi-Task Learning with Transformer EncodersHan He, Jinho D. ChoiEMNLP 2021 · 被引用 111 次
- Multi-Task Dense Prediction Fine-Tuning with Mixture of Fine-Grained ExpertsYangyang Xu, Xi Ye, Duo SuACM MM 2025
- EMT-NAS: Transferring architectural knowledge between tasks from different datasetsPeng Liao, Yaochu Jin, Wenli DuCVPR 2023
- Learning Conflict-Noticed Architecture for Multi-Task LearningZhixiong Yue, Yu Zhang, Jie LiangAAAI 2023 · 被引用 9 次
