xMTF: A Formula-Free Model for Reinforcement-Learning-Based Multi-Task Fusion in Recommender Systems
Yang Cao, Changhao Zhang, Xiaoshuang Chen, Kaiqiao Zhan, Ben Wang
Abstract
Recommender systems need to optimize various types of user feedback, e.g., clicks, likes, and shares. A typical recommender system handling multiple types of feedback has two components: a multitask learning (MTL) module, predicting feedback such as clickthrough rate and like rate; and a multi-task fusion (MTF) module, integrating these predictions into a single score for item ranking. MTF is essential for ensuring user satisfaction, as it directly influences recommendation outcomes. Recently, reinforcement learning (RL) has been applied to MTF tasks to improve long-term user satisfaction. However, existing RL-based MTF methods are formulabased methods, which only adjust limited coefficients within predefined formulas. The pre-defined formulas restrict the RL search space and become a bottleneck for MTF. To overcome this, we propose a formula-free MTF framework. We demonstrate that any suitable fusion function can be expressed as a composition of singlevariable monotonic functions, as per the Sprecher Representation Theorem. Leveraging this, we introduce a novel learnable monotonic fusion cell (MFC) to replace pre-defined formulas. We call this new MFC-based model eXtreme MTF (xMTF). Furthermore, we employ a two-stage hybrid (TSH) learning strategy to train xMTF effectively. By expanding the MTF search space, xMTF outperforms existing methods in extensive offline and online experiments. CCS Concepts • Information systems → Recommender systems.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 64194f78-09ae-4002-9db4-97d3e7000cf7Builds on7
- Gradient Surgery for Multi-Task LearningTianhe Yu, Saurabh Kumar, Abhishek Gupta, Sergey Levine et al.NeurIPS 2020 · 2,261 citations
- AdaTask: A Task-Aware Adaptive Learning Rate Approach to Multi-Task LearningEnneng Yang, Junwei Pan, Ximei Wang, Haibin Yu et al.AAAI 2023 · 70 citations
- Two-Stage Constrained Actor-Critic for Short Video RecommendationQingpeng Cai, Zhenghai Xue, Chi Zhang, Wanqi Xue et al.WWW 2023 · 60 citations
- PrefRec: Recommender Systems with Human Preferences for Reinforcing Long-term User EngagementWanqi Xue, Qingpeng Cai, Zhenghai Xue, Shuo Sun et al.KDD 2023 · 24 citations
- No Prior Mask: Eliminate Redundant Action for Deep Reinforcement LearningDianyu Zhong, Yiqin Yang, Qianchuan ZhaoAAAI 2024 · 15 citations
Related papers
- Multi-Task Recommendations with Reinforcement LearningZiru Liu, Jiejie Tian, Qingpeng Cai, Xiangyu Zhao et al.WWW 2023 · 57 citations
- Projection Optimization: A General Framework for Multi-Objective and Multi-Group RLHFNuoya Xiong, Aarti SinghICML 2025
- MT³: A Synergistic Multi-Task RL Framework for Specializing MLLMs in Text Image Machine TranslationZhaopeng Feng, Yupu Liang, Shaosheng Cao, Jiayuan Su et al.ACL 2026
- RMLer: Synthesizing Novel Objects Across Diverse Categories via Reinforcement Mixing LearningJun Li, Zikun Chen, Haibo Chen, Shuo Chen et al.AAAI 2026
- Who You Are Matters: Bridging Interests and Social Roles via LLM-Enhanced Logic RecommendationQing Yu, Xiaobei Wang, Shuchang Liu, Yandong Bai et al.NeurIPS 2025 · 4 citations
