Dynamic Knapsack Optimization Towards Efficient Multi-Channel Sequential Advertising
Xiaotian Hao, Zhaoqing Peng, Yi Ma, Guan Wang, Junqi Jin, Jianye Hao, Shan Chen, Rongquan Bai, Mingzhou Xie, Miao Xu, Zhenzhe Zheng, Chuan Yu
摘要
In E-commerce, advertising is essential for merchants to reach their target users. The typical objective is to maximize the advertiser's cumulative revenue over a period of time under a budget constraint. In real applications, an advertisement (ad) usually needs to be exposed to the same user multiple times until the user finally contributes revenue (e.g., places an order). However, existing advertising systems mainly focus on the immediate revenue with single ad exposures, ignoring the contribution of each exposure to the final conversion, thus usually falls into suboptimal solutions. In this paper, we formulate the sequential advertising strategy optimization as a dynamic knapsack problem. We propose a theoretically guaranteed bilevel optimization framework, which significantly reduces the solution space of the original optimization space while ensuring the solution quality. To improve the exploration efficiency of reinforcement learning, we also devise an effective action space reduction approach. Extensive offline and online experiments show the superior performance of our approaches over state-of-the-art baselines in terms of cumulative revenue.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- MetaDiffuser: Diffusion Model as Conditional Planner for Offline Meta-RLFei Ni, Jianye Hao, Yao Mu, Yifu Yuan 等ICML 2023 · 被引用 75 次
- Sustainable Online Reinforcement Learning for Auto-biddingZhiyu Mou, Yusen Huo, Rongquan Bai, Mingzhou Xie 等NeurIPS 2022 · 被引用 53 次
- Direct Heterogeneous Causal Learning for Resource Allocation Problems in MarketingHao Zhou, Shaoming Li, Guibin Jiang, Jiaqi Zheng 等AAAI 2023 · 被引用 35 次
- ERL-Re: Efficient Evolutionary Reinforcement Learning with Shared State Representation and Individual Policy RepresentationJianye Hao, Pengyi Li, Hongyao Tang, Yan Zheng 等ICLR 2023 · 被引用 16 次
- Iteratively Refined Behavior Regularization for Offline Reinforcement LearningYi Ma, Jianye Hao, Xiaohan Hu, Yan Zheng 等NeurIPS 2024 · 被引用 11 次
相关 Paper
- DEAR: Deep Reinforcement Learning for Online Advertising Impression in Recommender SystemsXiangyu Zhao, Changsheng Gu, Haoshenglun Zhang, Xiwang Yang 等AAAI 2021 · 被引用 131 次
- Maximizing the Success Probability of Policy Allocations in Online SystemsArtem Betlei, Mariia Vladimirova, Mehdi Sebbar, Nicolas Urien 等AAAI 2024 · 被引用 5 次
- SACO: Sequence-Aware Constrained Optimization Framework for Coupon Distribution in E-commerceLi Kong, Bingzhe Wang, Zhou Chen, Suhan Hu 等AAAI 2026
- MetaTrader: Learning to Generalize RL Trading Policies Beyond Offline DataHaochen Yuan, Minting Pan, Yunbo Wang, Siyu Gao 等AAAI 2026
- Auto-Bidding in Real-Time Auctions via Oracle Imitation LearningAlberto Silvio Chiappa, Briti Gangopadhyay, Zhao Wang, Shingo TakamatsuKDD 2025 · 被引用 4 次
