DEAR: Deep Reinforcement Learning for Online Advertising Impression in Recommender Systems
Xiangyu Zhao, Changsheng Gu, Haoshenglun Zhang, Xiwang Yang, Xiaobing Liu, Jiliang Tang, Hui Liu
Abstract
With the recent prevalence of Reinforcement Learning (RL), there have been tremendous interests in utilizing RL for online advertising in recommendation platforms (e.g., e-commerce and news feed sites). However, most RL-based advertising algorithms focus on optimizing ads' revenue while ignoring the possible negative influence of ads on user experience of recommended items (products, articles and videos). Developing an optimal advertising algorithm in recommendations faces immense challenges because interpolating ads improperly or too frequently may decrease user experience, while interpolating fewer ads will reduce the advertising revenue. Thus, in this paper, we propose a novel advertising strategy for the rec/ads trade-off. To be specific, we develop an RL-based framework that can continuously update its advertising strategies and maximize reward in the long run. Given a recommendation list, we design a novel Deep Q-network architecture that can determine three internally related tasks jointly, i.e., (i) whether to interpolate an ad or not in the recommendation list, and if yes, (ii) the optimal ad and (iii) the optimal location to interpolate. The experimental results based on real-world data demonstrate the effectiveness of the proposed framework.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9a7953aa-6308-4cce-8aff-99cd6f0b1372Cited by top-tier papers18
- LinRec: Linear Attention Mechanism for Long-term Sequential Recommender SystemsLangming Liu, Liu Cai, Chi Zhang, Xiangyu Zhao et al.SIGIR 2023 · 86 citations
- Process vs. Outcome Reward: Which is Better for Agentic RAG Reinforcement LearningWenlin Zhang, Xiangyang Li, Kuicai Dong, Yichao Wang et al.NeurIPS 2025 · 85 citations
- Fully Adaptive Framework: Neural Computerized Adaptive Testing for Online EducationYan Zhuang, Qi Liu, Zhenya Huang, Zhi Li et al.AAAI 2022 · 66 citations
- AutoDenoise: Automatic Data Instance Denoising for RecommendationsWeilin Lin, Xiangyu Zhao, Yejing Wang, Yuanshao Zhu et al.WWW 2023 · 62 citations
- Multi-Task Recommendations with Reinforcement LearningZiru Liu, Jiejie Tian, Qingpeng Cai, Xiangyu Zhao et al.WWW 2023 · 57 citations
Builds on2
Related papers
- On Designing the Optimal Integrated Ad Auction in E-commerce PlatformsYuchao Ma, Weian Li, Yuhan Wang, Zitian Guo et al.AAAI 2025
- Cross DQN: Cross Deep Q Network for Ads Allocation in FeedGuogang Liao, Ze Wang, Xiaoxu Wu, Xiaowen Shi et al.WWW 2022 · 46 citations
- An End-to-End Deep RL Framework for Task Arrangement in Crowdsourcing PlatformsCaihua Shan, Nikos Mamoulis, Reynold Cheng, Guoliang Li et al.ICDE 2020 · 23 citations
- Dynamic Knapsack Optimization Towards Efficient Multi-Channel Sequential AdvertisingXiaotian Hao, Zhaoqing Peng, Yi Ma, Guan Wang et al.ICML 2020 · 29 citations
- MaHRL: Multi-goals Abstraction Based Deep Hierarchical Reinforcement Learning for RecommendationsDongyang Zhao, Liang Zhang, Bo Zhang, Lizhou Zheng et al.SIGIR 2020 · 33 citations
