STEM: Unleashing the Power of Embeddings for Multi-Task Recommendation
Liangcai Su, Junwei Pan, Ximei Wang, Xi Xiao, Shijie Quan, Xihua Chen, Jie Jiang
Abstract
Multi-task learning (MTL) has gained significant popularity in recommender systems as it enables simultaneous optimization of multiple objectives. A key challenge in MTL is negative transfer, but existing studies explored negative transfer on all samples, overlooking the inherent complexities within them. We split the samples according to the relative amount of positive feedback among tasks. Surprisingly, negative transfer still occurs in existing MTL methods on samples that receive comparable feedback across tasks. Existing work commonly employs a shared-embedding paradigm, limiting the ability of modeling diverse user preferences on different tasks. In this paper, we introduce a novel Shared and Task-specific EMbeddings (STEM) paradigm that aims to incorporate both shared and task-specific embeddings to effectively capture taskspecific user preferences. Under this paradigm, we propose a simple model STEM-Net, which is equipped with an All Forward Task-specific Backward gating network to facilitate the learning of task-specific embeddings and direct knowledge transfer across tasks. Remarkably, STEM-Net demonstrates exceptional performance on comparable samples, achieving positive transfer. Comprehensive evaluation on three public MTL recommendation datasets demonstrates that STEM-Net outperforms state-of-the-art models by a substantial margin. Our code is released at https://github.com/LiangcaiSu/STEM .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers7
- PMG : Personalized Multimodal Generation with Large Language ModelsXiaoteng Shen, Rui Zhang, Xiaoyan Zhao, Jieming Zhu et al.WWW 2024 · 40 citations
- Pre-train, Align, and Disentangle: Empowering Sequential Recommendation with Large Language ModelsYuhao Wang, Junwei Pan, Pengyue Jia, Wanyu Wang et al.SIGIR 2025 · 8 citations
- Combinatorial Optimization Perspective based Framework for Multi-behavior RecommendationChenhao Zhai, Chang Meng, Yu Yang, Kexin Zhang et al.KDD 2025 · 4 citations
- xMTF: A Formula-Free Model for Reinforcement-Learning-Based Multi-Task Fusion in Recommender SystemsYang Cao, Changhao Zhang, Xiaoshuang Chen, Kaiqiao Zhan et al.WWW 2025 · 3 citations
- MultiTab: A Scalable Foundation for Multitask Learning on Tabular DataDimitrios Sinodinos, Jack Yi Wei, Narges ArmanfardAAAI 2026 · 2 citations
Builds on8
- Gradient Surgery for Multi-Task LearningTianhe Yu, Saurabh Kumar, Abhishek Gupta, Sergey Levine et al.NeurIPS 2020 · 2,261 citations
- DSelect-k: Differentiable Selection in the Mixture of Experts with Applications to Multi-Task LearningHussein Hazimeh, Zhe Zhao, Aakanksha Chowdhery, Maheswaran Sathiamoorthy et al.NeurIPS 2021 · 216 citations
- AdaTask: A Task-Aware Adaptive Learning Rate Approach to Multi-Task LearningEnneng Yang, Junwei Pan, Ximei Wang, Haibin Yu et al.AAAI 2023 · 70 citations
- MetaBalance: Improving Multi-Task Recommendations via Adapting Gradient Magnitudes of Auxiliary TasksYun He, Xue Feng, Cheng Cheng, Geng Ji et al.WWW 2022 · 69 citations
- Cross-Task Knowledge Distillation in Multi-Task RecommendationChenxiao Yang, Junwei Pan, Xiaofeng Gao, Tingyu Jiang et al.AAAI 2022 · 58 citations
Related papers
- A Contrastive Sharing Model for Multi-Task RecommendationTing Bai, Yudong Xiao, Bin Wu, Guojun Yang et al.WWW 2022 · 30 citations
- Direct Routing Gradient (DRGrad): A Personalized Information Surgery for Multi-Task Learning (MTL) RecommendationsYuguang Liu, Yiyun Miao, Luyao XiaAAAI 2025 · 2 citations
- Selective Task Group Updates for Multi-Task OptimizationWooseong Jeong, Kuk-Jin YoonICLR 2025
- Automatic Multi-Task Learning Framework with Neural Architecture Search in RecommendationsShen Jiang, Guanghui Zhu, Yue Wang, Chunfeng Yuan et al.KDD 2024 · 5 citations
- Towards Consistent Multi-Task Learning: Unlocking the Potential of Task-Specific ParametersXiaohan Qin, Xiaoxing Wang, Junchi YanCVPR 2025
