On the Unexpected Effectiveness of Reinforcement Learning for Sequential Recommendation
Alvaro Labarca, Denis Parra, Rodrigo Toro Icarte
摘要
In recent years, Reinforcement Learning (RL) has shown great promise in session-based recommendation. Sequential models that use RL have reached state-of-the-art performance for the Next-item Prediction (NIP) task. This result is intriguing, as the NIP task only evaluates how well the system can correctly recommend the next item to the user, while the goal of RL is to find a policy that optimizes rewards in the long term -sometimes at the expense of suboptimal shortterm performance. Then, how can RL improve the system's performance on short-term metrics? This article investigates this question by exploring proxy learning objectives, which we identify as goals RL models might be following, and thus could explain the performance boost. We found that RL -when used as an auxiliary loss -promotes the learning of embeddings that capture information about the user's previously interacted items. Subsequently, we replaced the RL objective with a straightforward auxiliary loss designed to predict the number of items the user interacted with. This substitution results in performance gains comparable to RL. These findings pave the way to improve performance and understanding of RL methods for recommender systems.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper5
- Self-Supervised Reinforcement Learning for Recommender SystemsXin Xin, Alexandros Karatzoglou, Ioannis Arapakis, Joemon M. JoseSIGIR 2020 · 被引用 217 次
- Leveraging Demonstrations for Reinforcement Recommendation Reasoning over Knowledge GraphsKangzhi Zhao, Xiting Wang, Yuren Zhang, Li Zhao 等SIGIR 2020 · 被引用 114 次
- A General Offline Reinforcement Learning Framework for Interactive RecommendationTeng Xiao, Donglin WangAAAI 2021 · 被引用 82 次
- Reinforced Anchor Knowledge Graph Generation for News Recommendation ReasoningDanyang Liu, Jianxun Lian, Zheng Liu, Xiting Wang 等KDD 2021 · 被引用 49 次
- Rethinking Reinforcement Learning for Recommendation: A Prompt PerspectiveXin Xin, Tiago Pimentel, Alexandros Karatzoglou, Pengjie Ren 等SIGIR 2022 · 被引用 48 次
相关 Paper
- Unsupervised Proxy Selection for Session-based Recommender SystemsJunsu Cho, SeongKu Kang, Dongmin Hyun, Hwanjo YuSIGIR 2021 · 被引用 21 次
- KERL: A Knowledge-Guided Reinforcement Learning Model for Sequential RecommendationPengfei Wang, Yu Fan, Long Xia, Wayne Xin Zhao 等SIGIR 2020 · 被引用 122 次
- Incorporating User Micro-behaviors and Item Knowledge into Multi-task Learning for Session-based RecommendationWenjing Meng, Deqing Yang, Yanghua XiaoSIGIR 2020 · 被引用 122 次
- ResAct: Reinforcing Long-term Engagement in Sequential Recommendation with Residual ActorWanqi Xue, Qingpeng Cai, Ruohan Zhan, Dong Zheng 等ICLR 2023 · 被引用 6 次
- NP-MiSR: Neural Process-based Multi-Interest Learning for Session-Based RecommendationJun Bao, Junbo Wang, Yiheng Jiang, Xiangfeng Liu 等AAAI 2026
