Controllable Dynamic Multi-Task Architectures
Dripta S. Raychaudhuri, Yumin Suh, Samuel Schulter, Xiang Yu, Masoud Faraki, Amit K. Roy-Chowdhury, Manmohan Chandraker
摘要
Multi-task learning commonly encounters competition for resources among tasks, specifically when model capac-ity is limited. This challenge motivates models which al-low control over the relative importance of tasks and total compute cost during inference time. In this work, we pro-pose such a controllable multi-task network that dynami-cally adjusts its architecture and weights to match the de-sired task preference as well as the resource constraints. In contrast to the existing dynamic multi-task approaches that adjust only the weights within a fixed architecture, our approach affords the flexibility to dynamically control the total computational cost and match the user-preferred task importance better. We propose a disentangled training of two hype rnetwo rks, by exploiting task affinity and a novel branching regularized loss, to take input prefer-ences and accordingly predict tree-structured models with adapted weights. Experiments on three multi-task bench-marks, namely PASCAL-Context, NYU-v2, and CIFAR-100, show the efficacy of our approach. Project page is available at https://www.nec-labs.com/-mas/DYMU.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper17
- Hypervolume Maximization: A Geometric View of Pareto Set LearningXiaoyuan Zhang, Xi Lin, Bo Xue, Yifan Chen 等NeurIPS 2023 · 被引用 40 次
- Multi-task Graph Neural Architecture Search with Task-aware Collaboration and CurriculumYijian Qin, Xin Wang, Ziwei Zhang, Hong Chen 等NeurIPS 2023 · 被引用 27 次
- Multi-Task Learning with Knowledge Distillation for Dense PredictionYangyang Xu, Yibo Yang, Lefei ZhangICCV 2023 · 被引用 18 次
- Efficient Controllable Multi-Task ArchitecturesAbhishek Aich, Samuel Schulter, Amit K. Roy-Chowdhury, Manmohan Chandraker 等ICCV 2023 · 被引用 9 次
- Growing a Brain with Sparsity-Inducing Generation for Continual LearningHyundong Jin, Gyeong-Hyeon Kim, Chanho Ahn, Eunwoo KimICCV 2023 · 被引用 7 次
它引用的顶会 Paper13
- Gradient Surgery for Multi-Task LearningTianhe Yu, Saurabh Kumar, Abhishek Gupta, Sergey Levine 等NeurIPS 2020 · 被引用 2,261 次
- Tent: Fully Test-Time Adaptation by Entropy MinimizationDequan Wang, Evan Shelhamer, Shaoteng Liu, Bruno A. Olshausen 等ICLR 2021 · 被引用 1,731 次
- AdaShare: Learning What To Share For Efficient Deep Multi-Task LearningXimeng Sun, Rameswar Panda, Rogério Feris, Kate SaenkoNeurIPS 2020 · 被引用 337 次
- Just Pick a Sign: Optimizing Deep Multitask Models with Gradient Sign DropoutZhao Chen, Jiquan Ngiam, Yanping Huang, Thang Luong 等NeurIPS 2020 · 被引用 313 次
- Learning to Branch for Multi-Task LearningPengsheng Guo, Chen-Yu Lee, Daniel UlbrichtICML 2020 · 被引用 208 次
相关 Paper
- Task Switching Network for Multi-task LearningGuolei Sun, Thomas Probst, Danda Pani Paudel, Nikola Popovic 等ICCV 2021 · 被引用 57 次
- Deep Elastic Networks With Model Selection for Multi-Task LearningChanho Ahn, Eunwoo Kim, Songhwai OhICCV 2019 · 被引用 56 次
- DiSparse: Disentangled Sparsification for Multitask Model CompressionXinglong Sun, Ali Hassani, Zhangyang Wang, Gao Huang 等CVPR 2022 · 被引用 22 次
- Which Tasks Should Be Learned Together in Multi-task Learning?Trevor Standley, Amir Zamir, Dawn Chen, Leonidas J. Guibas 等ICML 2020 · 被引用 651 次
- Learning with Privileged TasksYuru Song, Zan Lou, Shan You, Erkun Yang 等ICCV 2021 · 被引用 3 次
