MetaBalance: Improving Multi-Task Recommendations via Adapting Gradient Magnitudes of Auxiliary Tasks
Yun He, Xue Feng, Cheng Cheng, Geng Ji, Yunsong Guo, James Caverlee
摘要
In many personalized recommendation scenarios, the generalization ability of a target task can be improved via learning with additional auxiliary tasks alongside this target task on a multi-task network. However, this method often suffers from a serious optimization imbalance problem. On the one hand, one or more auxiliary tasks might have a larger influence than the target task and even dominate the network weights, resulting in worse recommendation accuracy for the target task. On the other hand, the influence of one or more auxiliary tasks might be too weak to assist the target task. More challenging is that this imbalance dynamically changes throughout the training process and varies across the parts of the same network. We propose a new method: MetaBalance to balance auxiliary losses via directly manipulating their gradients w.r.t the shared parameters in the multi-task network. Specifically, in each training iteration and adaptively for each part of the network, the gradient of an auxiliary loss is carefully reduced or enlarged to have a closer magnitude to the gradient of the target loss, preventing auxiliary tasks from being so strong that dominate the target task or too weak to help the target task. Moreover, the proximity between the gradient magnitudes can be flexibly adjusted to adapt MetaBalance to different scenarios. The experiments show that our proposed method achieves a significant improvement of 8.34% in terms of NDCG@10 upon the strongest baseline on two real-world datasets. The code of our approach can be found at here.1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper15
- AdaMerging: Adaptive Model Merging for Multi-Task LearningEnneng Yang, Zhenyi Wang, Li Shen, Shiwei Liu 等ICLR 2024 · 被引用 230 次
- MMPareto: Boosting Multimodal Learning with Innocent Unimodal AssistanceYake Wei, Di HuICML 2024 · 被引用 86 次
- Multi-behavior Self-supervised Learning for RecommendationJingcao Xu, Chaokun Wang, Cheng Wu, Yang Song 等SIGIR 2023 · 被引用 80 次
- AdaTask: A Task-Aware Adaptive Learning Rate Approach to Multi-Task LearningEnneng Yang, Junwei Pan, Ximei Wang, Haibin Yu 等AAAI 2023 · 被引用 70 次
- Smooth Tchebycheff Scalarization for Multi-Objective OptimizationXi Lin, Xiaoyuan Zhang, Zhiyuan Yang, Fei Liu 等ICML 2024 · 被引用 48 次
它引用的顶会 Paper3
- Gradient Surgery for Multi-Task LearningTianhe Yu, Saurabh Kumar, Abhishek Gupta, Sergey Levine 等NeurIPS 2020 · 被引用 2,261 次
- Just Pick a Sign: Optimizing Deep Multitask Models with Gradient Sign DropoutZhao Chen, Jiquan Ngiam, Yanping Huang, Thang Luong 等NeurIPS 2020 · 被引用 313 次
- MTAdam: Automatic Balancing of Multiple Training Loss TermsItzik Malkiel, Lior WolfEMNLP 2021 · 被引用 12 次
相关 Paper
- Towards Impartial Multi-task LearningLiyang Liu, Yi Li, Zhanghui Kuang, Jing-Hao Xue 等ICLR 2021 · 被引用 228 次
- Can Small Heads Help? Understanding and Improving Multi-Task GeneralizationYuyan Wang, Zhe Zhao, Bo Dai, Christopher Fifty 等WWW 2022 · 被引用 15 次
- Learning with Privileged TasksYuru Song, Zan Lou, Shan You, Erkun Yang 等ICCV 2021 · 被引用 3 次
- Auxiliary Learning as an Asymmetric Bargaining GameAviv Shamsian, Aviv Navon, Neta Glazer, Kenji Kawaguchi 等ICML 2023 · 被引用 15 次
- Quantifying Task Priority for Multi-Task OptimizationWooseong Jeong, Kuk-Jin YoonCVPR 2024
