Contrastive State Augmentations for Reinforcement Learning-Based Recommender Systems
Zhaochun Ren, Na Huang, Yidan Wang, Pengjie Ren, Jun Ma, Jiahuan Lei, Xinlei Shi, Hengliang Luo, Joemon M. Jose, Xin Xin
摘要
Learning reinforcement learning (RL)-based recommenders from historical user-item interaction sequences is vital to generate highreward recommendations and improve long-term cumulative benefits. However, existing RL recommendation methods encounter difficulties (i) to estimate the value functions for states which are not contained in the offline training data, and (ii) to learn effective state representations from user implicit feedback due to the lack of contrastive signals.
In this work, we propose contrastive state augmentations (CSA) for the training of RL-based recommender systems. To tackle the first issue, we propose four state augmentation strategies to enlarge the state space of the offline data. The proposed method improves the generalization capability of the recommender by making the RL agent visit the local state regions and ensuring the learned value functions are similar between the original and augmented states. For the second issue, we propose introducing contrastive signals between augmented states and the state randomly sampled from other sessions to improve the state representation learning further.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Reinforcement Learning-based Recommender Systems with Large Language Models for State Reward and Action ModelingJie Wang, Alexandros Karatzoglou, Ioannis Arapakis, Joemon M. JoseSIGIR 2024 · 被引用 27 次
- Data Augmentation as Free Lunch: Exploring the Test-Time Augmentation for Sequential RecommendationYizhou Dang, Yuting Liu, Enneng Yang, Minhan Huang 等SIGIR 2025 · 被引用 10 次
- A Generalised and Adaptable Reinforcement Learning Stopping MethodReem Bin Hezam, Mark StevensonSIGIR 2025 · 被引用 1 次
它引用的顶会 Paper11
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- Conservative Q-Learning for Offline Reinforcement LearningAviral Kumar, Aurick Zhou, George Tucker, Sergey LevineNeurIPS 2020 · 被引用 2,881 次
- SimCSE: Simple Contrastive Learning of Sentence EmbeddingsTianyu Gao, Xingcheng Yao, Danqi ChenEMNLP 2021 · 被引用 2,496 次
- Self-supervised Graph Learning for RecommendationJiancan Wu, Xiang Wang, Fuli Feng, Xiangnan He 等SIGIR 2021 · 被引用 1,476 次
- Self-Supervised Hypergraph Convolutional Networks for Session-based RecommendationXin Xia, Hongzhi Yin, Junliang Yu, Qinyong Wang 等AAAI 2021 · 被引用 615 次
相关 Paper
- Rethinking Reinforcement Learning for Recommendation: A Prompt PerspectiveXin Xin, Tiago Pimentel, Alexandros Karatzoglou, Pengjie Ren 等SIGIR 2022 · 被引用 48 次
- Is Contrastive Learning Necessary? A Study of Data Augmentation vs Contrastive Learning in Sequential RecommendationPeilin Zhou, You-Liang Huang, Yueqi Xie, Jingqi Gao 等WWW 2024 · 被引用 36 次
- Contrastive Learning for Sequential RecommendationXu Xie, Fei Sun, Zhaoyang Liu, Shiwen Wu 等ICDE 2022 · 被引用 674 次
- Contrastive Representation for Interactive RecommendationJingyu Li, Zhiyong Feng, Dongxiao He, Hongqi Chen 等AAAI 2025 · 被引用 2 次
- Self-Supervised Reinforcement Learning for Recommender SystemsXin Xin, Alexandros Karatzoglou, Ioannis Arapakis, Joemon M. JoseSIGIR 2020 · 被引用 217 次
