Maximizing the Success Probability of Policy Allocations in Online Systems
Artem Betlei, Mariia Vladimirova, Mehdi Sebbar, Nicolas Urien, Thibaud Rahier, Benjamin Heymann
Abstract
The effectiveness of advertising in e-commerce largely depends on the ability of merchants to bid on and win impressions for their targeted users. The bidding procedure is highly complex due to various factors such as market competition, user behavior, and the diverse objectives of advertisers. In this paper we consider the problem at the level of user timelines instead of individual bid requests, manipulating full policies (i.e. pre-defined bidding strategies) and not bid values. In order to optimally allocate policies to users, typical multiple treatments allocation methods solve knapsack-like problems which aim at maximizing an expected value under constraints. In the industrial contexts such as online advertising, we argue that optimizing for the probability of success is a more suited objective than expected value maximization, and we introduce the SuccessProbaMax algorithm that aims at finding the policy allocation which is the most likely to outperform a fixed reference policy. Finally, we conduct comprehensive experiments both on synthetic and real-world data to evaluate its performance. The results demonstrate that our proposed algorithm outperforms conventional expectedvalue maximization algorithms in terms of success rate.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on4
- A Unifying Framework for Online Optimization with Long-Term ConstraintsMatteo Castiglioni, Andrea Celli, Alberto Marchesi, Giulia Romano et al.NeurIPS 2022 · 59 citations
- LBCF: A Large-Scale Budget-Constrained Causal Forest AlgorithmMeng Ai, Biao Li, Heyang Gong, Qingwei Yu et al.WWW 2022 · 27 citations
- Personalized Treatment Selection using Causal HeterogeneityYe Tu, Kinjal Basu, Cyrus DiCiccio, Romil Bansal et al.WWW 2021 · 11 citations
- Causal Models for Real Time Bidding with Repeated User InteractionsMartin Bompaire, Alexandre Gilotte, Benjamin HeymannKDD 2021 · 10 citations
Related papers
- Dynamic Knapsack Optimization Towards Efficient Multi-Channel Sequential AdvertisingXiaotian Hao, Zhaoqing Peng, Yi Ma, Guan Wang et al.ICML 2020 · 29 citations
- Stochastic bandits for multi-platform budget optimization in online advertisingVashist Avadhanula, Riccardo Colini-Baldeschi, Stefano Leonardi, Karthik Abinav Sankararaman et al.WWW 2021 · 43 citations
- Optimising Budget Management via Primal-Dual Approximation with Constrained Polynomial Weights UpdateDmitrii Moor, Per Berglund, Hannes Karlbom, Zhenwen Dai et al.KDD 2025
- No-Regret Algorithms in non-Truthful Auctions with Budget and ROI ConstraintsGagan Aggarwal, Giannis Fikioris, Mingfei ZhaoWWW 2025 · 13 citations
- Dual Mirror Descent for Online Allocation ProblemsSantiago R. Balseiro, Haihao Lu, Vahab S. MirrokniICML 2020 · 102 citations
