Hierarchical Reinforcement Learning for Integrated Recommendation
Ruobing Xie, Shaoliang Zhang, Rui Wang, Feng Xia, Leyu Lin
Abstract
Integrated recommendation aims to jointly recommend heterogeneous items in the main feed from different sources via multiple channels, which needs to capture user preferences on both item and channel levels. It has been widely used in practical systems by billions of users, while few works concentrate on the integrated recommendation systematically. In this work, we propose a novel Hierarchical reinforcement learning framework for integrated recommendation (HRL-Rec), which divides the integrated recommendation into two tasks to recommend channels and items sequentially. The low-level agent is a channel selector, which generates a personalized channel list. The high-level agent is an item recommender, which recommends specific items from heterogeneous channels under the channel constraints. We design various rewards for both recommendation accuracy and diversity, and propose four losses for fast and stable model convergence. We also conduct an online exploration for sufficient training. In experiments, we conduct extensive offline and online experiments on a billion-level real-world dataset to show the effectiveness of HRL-Rec. HRL-Rec has also been deployed on WeChat Top Stories, affecting millions of users. The source codes are released in https://github.com/modriczhang/HRL-Rec . * indicates equal contribution. Ruobing Xie is the corresponding author (
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 034e337a-e17f-40e7-a364-48436ef53561Cited by top-tier papers12
- Curriculum Disentangled Recommendation with Noisy Multi-feedbackHong Chen, Yudong Chen, Xin Wang, Ruobing Xie et al.NeurIPS 2021 · 88 citations
- Hierarchical Diffusion for Offline Decision MakingWenhao Li, Xiangfeng Wang, Bo Jin, Hongyuan ZhaICML 2023 · 80 citations
- Cross DQN: Cross Deep Q Network for Ads Allocation in FeedGuogang Liao, Ze Wang, Xiaoxu Wu, Xiaowen Shi et al.WWW 2022 · 46 citations
- Personalized Approximate Pareto-Efficient RecommendationRuobing Xie, Yanlei Liu, Shaoliang Zhang, Rui Wang et al.WWW 2021 · 46 citations
- Controlling Underestimation Bias in Reinforcement Learning via Quasi-median OperationWei Wei, Yujia Zhang, Jiye Liang, Lin Li et al.AAAI 2022 · 20 citations
Builds on3
- Self-Supervised Reinforcement Learning for Recommender SystemsXin Xin, Alexandros Karatzoglou, Ioannis Arapakis, Joemon M. JoseSIGIR 2020 · 217 citations
- Adaptive Factorization Network: Learning Adaptive-Order Feature InteractionsWeiyu Cheng, Yanyan Shen, Linpeng HuangAAAI 2020 · 202 citations
- MaHRL: Multi-goals Abstraction Based Deep Hierarchical Reinforcement Learning for RecommendationsDongyang Zhao, Liang Zhang, Bo Zhang, Lizhou Zheng et al.SIGIR 2020 · 33 citations
Related papers
- DEAR: Deep Reinforcement Learning for Online Advertising Impression in Recommender SystemsXiangyu Zhao, Changsheng Gu, Haoshenglun Zhang, Xiwang Yang et al.AAAI 2021 · 131 citations
- Deep Unified Representation for Heterogeneous RecommendationChengqiang Lu, Mingyang Yin, Shuheng Shen, Luo Ji et al.WWW 2022 · 9 citations
- HieRec: Hierarchical User Interest Modeling for Personalized News RecommendationTao Qi, Fangzhao Wu, Chuhan Wu, Peiru Yang et al.ACL 2021
- Dual Contrastive Transformer for Hierarchical Preference Modeling in Sequential RecommendationChengkai Huang, Shoujin Wang, Xianzhi Wang, Lina YaoSIGIR 2023 · 17 citations
- Package Recommendation with Intra- and Inter-Package Attention NetworksChen Li, Yuanfu Lu, Wei Wang, Chuan Shi et al.SIGIR 2021 · 15 citations
