Cross DQN: Cross Deep Q Network for Ads Allocation in Feed
Guogang Liao, Ze Wang, Xiaoxu Wu, Xiaowen Shi, Chuheng Zhang, Yongkang Wang, Xingxing Wang, Dong Wang
Abstract
E-commerce platforms usually display a mixed list of ads and organic items in feed. One key problem is to allocate the limited slots in the feed to maximize the overall revenue as well as improve user experience, which requires a good model for user preference. Instead of modeling the influence of individual items on user behaviors, the arrangement signal models the influence of the arrangement of items and may lead to a better allocation strategy. However, most of previous strategies fail to model such a signal and therefore result in suboptimal performance. In addition, the percentage of ads exposed (PAE) is an important indicator in ads allocation. Excessive PAE hurts user experience while too low PAE reduces platform revenue. Therefore, how to constrain the PAE within a certain range while keeping personalized recommendation under the PAE constraint is a challenge. In this paper, we propose Cross Deep Q Network (Cross DQN) to extract the crucial arrangement signal by crossing the embeddings of different items and modeling the crossed sequence by multi-channel attention. Besides, we propose an auxiliary loss for batch-level constraint on PAE to tackle the above-mentioned challenge. Our model results in higher revenue and better user experience than state-of-the-art baselines in offline experiments. Moreover, our model demonstrates a significant improvement in the online A/B test and has been fully deployed on Meituan feed to serve more than 300 millions of customers.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext bebd6adb-9c55-4de7-9461-93b23a6e7af0Cited by top-tier papers7
- RL-MPCA: A Reinforcement Learning Based Multi-Phase Computation Allocation Approach for Recommender SystemsJiahong Zhou, Shunhui Mao, Guoliang Yang, Bo Tang et al.WWW 2023 · 10 citations
- Whittle Index with Multiple Actions and State Constraint for Inventory ManagementChuheng Zhang, Xiangsen Wang, Wei Jiang, Xianliang Yang et al.ICLR 2024 · 10 citations
- Deep Automated Mechanism Design for Integrating Ad Auction and Allocation in FeedXuejian Li, Ze Wang, Bingqi Zhu, Fei He et al.SIGIR 2024 · 9 citations
- DeCoCDR: Deployable Cloud-Device Collaboration for Cross-Domain RecommendationYu Li, Yi Zhang, Zimu Zhou, Qiang LiSIGIR 2024 · 2 citations
- On Designing the Optimal Integrated Ad Auction in E-commerce PlatformsYuchao Ma, Weian Li, Yuhan Wang, Zitian Guo et al.AAAI 2025
Builds on2
- DEAR: Deep Reinforcement Learning for Online Advertising Impression in Recommender SystemsXiangyu Zhao, Changsheng Gu, Haoshenglun Zhang, Xiwang Yang et al.AAAI 2021 · 131 citations
- Hierarchical Reinforcement Learning for Integrated RecommendationRuobing Xie, Shaoliang Zhang, Rui Wang, Feng Xia et al.AAAI 2021 · 90 citations
Related papers
- A Context-Aware Framework for Integrating Ad Auctions and RecommendationsYuchao Ma, Weian Li, Yuejia Dou, Zhiyuan Su et al.WWW 2025 · 3 citations
- Decision-Making Context Interaction Network for Click-Through Rate PredictionXiang Li, Shuwei Chen, Jian Dong, Jin Zhang et al.AAAI 2023 · 11 citations
- An End-to-End Deep RL Framework for Task Arrangement in Crowdsourcing PlatformsCaihua Shan, Nikos Mamoulis, Reynold Cheng, Guoliang Li et al.ICDE 2020 · 23 citations
- Efficient and Practical Approximation Algorithms for Advertising in Content FeedsGuangyi Zhang, Ilie Sarpe, Aristides GionisWWW 2025
- Beyond Advertising: Mechanism Design for Platform-Wide Marketing Service "QuanZhanTui"Ningyuan Li, Zhilin Zhang, Tianyan Long, Yuyao Liu et al.KDD 2025 · 1 citation
