Cross DQN: Cross Deep Q Network for Ads Allocation in Feed
Guogang Liao, Ze Wang, Xiaoxu Wu, Xiaowen Shi, Chuheng Zhang, Yongkang Wang, Xingxing Wang, Dong Wang
摘要
E-commerce platforms usually display a mixed list of ads and organic items in feed. One key problem is to allocate the limited slots in the feed to maximize the overall revenue as well as improve user experience, which requires a good model for user preference. Instead of modeling the influence of individual items on user behaviors, the arrangement signal models the influence of the arrangement of items and may lead to a better allocation strategy. However, most of previous strategies fail to model such a signal and therefore result in suboptimal performance. In addition, the percentage of ads exposed (PAE) is an important indicator in ads allocation. Excessive PAE hurts user experience while too low PAE reduces platform revenue. Therefore, how to constrain the PAE within a certain range while keeping personalized recommendation under the PAE constraint is a challenge. In this paper, we propose Cross Deep Q Network (Cross DQN) to extract the crucial arrangement signal by crossing the embeddings of different items and modeling the crossed sequence by multi-channel attention. Besides, we propose an auxiliary loss for batch-level constraint on PAE to tackle the above-mentioned challenge. Our model results in higher revenue and better user experience than state-of-the-art baselines in offline experiments. Moreover, our model demonstrates a significant improvement in the online A/B test and has been fully deployed on Meituan feed to serve more than 300 millions of customers.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- RL-MPCA: A Reinforcement Learning Based Multi-Phase Computation Allocation Approach for Recommender SystemsJiahong Zhou, Shunhui Mao, Guoliang Yang, Bo Tang 等WWW 2023 · 被引用 10 次
- Whittle Index with Multiple Actions and State Constraint for Inventory ManagementChuheng Zhang, Xiangsen Wang, Wei Jiang, Xianliang Yang 等ICLR 2024 · 被引用 10 次
- Deep Automated Mechanism Design for Integrating Ad Auction and Allocation in FeedXuejian Li, Ze Wang, Bingqi Zhu, Fei He 等SIGIR 2024 · 被引用 9 次
- DeCoCDR: Deployable Cloud-Device Collaboration for Cross-Domain RecommendationYu Li, Yi Zhang, Zimu Zhou, Qiang LiSIGIR 2024 · 被引用 2 次
- On Designing the Optimal Integrated Ad Auction in E-commerce PlatformsYuchao Ma, Weian Li, Yuhan Wang, Zitian Guo 等AAAI 2025
它引用的顶会 Paper2
- DEAR: Deep Reinforcement Learning for Online Advertising Impression in Recommender SystemsXiangyu Zhao, Changsheng Gu, Haoshenglun Zhang, Xiwang Yang 等AAAI 2021 · 被引用 131 次
- Hierarchical Reinforcement Learning for Integrated RecommendationRuobing Xie, Shaoliang Zhang, Rui Wang, Feng Xia 等AAAI 2021 · 被引用 90 次
相关 Paper
- A Context-Aware Framework for Integrating Ad Auctions and RecommendationsYuchao Ma, Weian Li, Yuejia Dou, Zhiyuan Su 等WWW 2025 · 被引用 3 次
- Decision-Making Context Interaction Network for Click-Through Rate PredictionXiang Li, Shuwei Chen, Jian Dong, Jin Zhang 等AAAI 2023 · 被引用 11 次
- An End-to-End Deep RL Framework for Task Arrangement in Crowdsourcing PlatformsCaihua Shan, Nikos Mamoulis, Reynold Cheng, Guoliang Li 等ICDE 2020 · 被引用 23 次
- Efficient and Practical Approximation Algorithms for Advertising in Content FeedsGuangyi Zhang, Ilie Sarpe, Aristides GionisWWW 2025
- Beyond Advertising: Mechanism Design for Platform-Wide Marketing Service "QuanZhanTui"Ningyuan Li, Zhilin Zhang, Tianyan Long, Yuyao Liu 等KDD 2025 · 被引用 1 次
