DEAR: Deep Reinforcement Learning for Online Advertising Impression in Recommender Systems
Xiangyu Zhao, Changsheng Gu, Haoshenglun Zhang, Xiwang Yang, Xiaobing Liu, Jiliang Tang, Hui Liu
摘要
With the recent prevalence of Reinforcement Learning (RL), there have been tremendous interests in utilizing RL for online advertising in recommendation platforms (e.g., e-commerce and news feed sites). However, most RL-based advertising algorithms focus on optimizing ads' revenue while ignoring the possible negative influence of ads on user experience of recommended items (products, articles and videos). Developing an optimal advertising algorithm in recommendations faces immense challenges because interpolating ads improperly or too frequently may decrease user experience, while interpolating fewer ads will reduce the advertising revenue. Thus, in this paper, we propose a novel advertising strategy for the rec/ads trade-off. To be specific, we develop an RL-based framework that can continuously update its advertising strategies and maximize reward in the long run. Given a recommendation list, we design a novel Deep Q-network architecture that can determine three internally related tasks jointly, i.e., (i) whether to interpolate an ad or not in the recommendation list, and if yes, (ii) the optimal ad and (iii) the optimal location to interpolate. The experimental results based on real-world data demonstrate the effectiveness of the proposed framework.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper18
- LinRec: Linear Attention Mechanism for Long-term Sequential Recommender SystemsLangming Liu, Liu Cai, Chi Zhang, Xiangyu Zhao 等SIGIR 2023 · 被引用 86 次
- Process vs. Outcome Reward: Which is Better for Agentic RAG Reinforcement LearningWenlin Zhang, Xiangyang Li, Kuicai Dong, Yichao Wang 等NeurIPS 2025 · 被引用 85 次
- Fully Adaptive Framework: Neural Computerized Adaptive Testing for Online EducationYan Zhuang, Qi Liu, Zhenya Huang, Zhi Li 等AAAI 2022 · 被引用 66 次
- AutoDenoise: Automatic Data Instance Denoising for RecommendationsWeilin Lin, Xiangyu Zhao, Yejing Wang, Yuanshao Zhu 等WWW 2023 · 被引用 62 次
- Multi-Task Recommendations with Reinforcement LearningZiru Liu, Jiejie Tian, Qingpeng Cai, Xiangyu Zhao 等WWW 2023 · 被引用 57 次
它引用的顶会 Paper2
相关 Paper
- On Designing the Optimal Integrated Ad Auction in E-commerce PlatformsYuchao Ma, Weian Li, Yuhan Wang, Zitian Guo 等AAAI 2025
- Cross DQN: Cross Deep Q Network for Ads Allocation in FeedGuogang Liao, Ze Wang, Xiaoxu Wu, Xiaowen Shi 等WWW 2022 · 被引用 46 次
- An End-to-End Deep RL Framework for Task Arrangement in Crowdsourcing PlatformsCaihua Shan, Nikos Mamoulis, Reynold Cheng, Guoliang Li 等ICDE 2020 · 被引用 23 次
- Dynamic Knapsack Optimization Towards Efficient Multi-Channel Sequential AdvertisingXiaotian Hao, Zhaoqing Peng, Yi Ma, Guan Wang 等ICML 2020 · 被引用 29 次
- MaHRL: Multi-goals Abstraction Based Deep Hierarchical Reinforcement Learning for RecommendationsDongyang Zhao, Liang Zhang, Bo Zhang, Lizhou Zheng 等SIGIR 2020 · 被引用 33 次
