The Limits of Optimal Pricing in the Dark
Quinlan Dawkins, Minbiao Han, Haifeng Xu
Abstract
A ubiquitous learning problem in today's digital market is, during repeated interactions between a seller and a buyer, how a seller can gradually learn optimal pricing decisions based on the buyer's past purchase responses. A fundamental challenge of learning in such a strategic setup is that the buyer will naturally have incentives to manipulate his responses in order to induce more favorable learning outcomes for him. To understand the limits of the seller's learning when facing such a strategic and possibly manipulative buyer, we study a natural yet powerful buyer manipulation strategy. That is, before the pricing game starts, the buyer simply commits to "imitate" a different value function by pretending to always react optimally according to this imitative value function. We fully characterize the optimal imitative value function that the buyer should imitate as well as the resultant seller revenue and buyer surplus under this optimal buyer manipulation. Our characterizations reveal many useful insights about what happens at equilibrium. For example, a seller with concave production cost will obtain essentially 0 revenue at equilibrium whereas the revenue for a seller with convex production cost is the Bregman divergence of her cost function between no production and certain production. Finally, and importantly, we show that a more powerful class of pricing schemes does not necessarily increase, in fact, may be harmful to, the seller's revenue. Our results not only lead to an effective prescriptive way for buyers to manipulate learning algorithms but also shed lights on the limits of what a seller can really achieve when pricing in the dark.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4486c08d-3339-424d-8afa-eff1e91d7293Cited by top-tier papers3
- Learning in Online Principal-Agent Interactions: The Power of MenusMinbiao Han, Michael Albert, Haifeng XuAAAI 2024 · 9 citations
- First-Order Convex Fitting and Its Application to Economics and OptimizationQuinlan Dawkins, Minbiao Han, Haifeng XuAAAI 2022 · 4 citations
- Optimal Pricing for Data-Augmented AutoML MarketplacesMinbiao Han, Steven Xia, Jonathan Li, Raul Castro Fernandez et al.ICML 2026 · 2 citations
Builds on4
- Learning Strategy-Aware Linear ClassifiersYiling Chen, Yang Liu, Chara PodimataNeurIPS 2020 · 110 citations
- Incentive-Aware PAC LearningHanrui Zhang, Vincent ConitzerAAAI 2021 · 54 citations
- Optimally Deceiving a Learning Leader in Stackelberg GamesGeorgios Birmpas, Jiarui Gan, Alexandros Hollender, Francisco J. Marmolejo Cossío et al.NeurIPS 2020 · 25 citations
- Learning the Valuations of a k-demand AgentHanrui Zhang, Vincent ConitzerICML 2020 · 10 citations
Related papers
- Bisection-Based Pricing for Repeated Contextual Auctions against Strategic BuyerAnton Zhiyanov, Alexey DrutsaICML 2020 · 11 citations
- Reserve Pricing in Repeated Second-Price Auctions with Strategic BiddersAlexey DrutsaICML 2020 · 17 citations
- Dynamic Pricing and Learning with Bayesian PersuasionShipra Agrawal, Yiding Feng, Wei TangNeurIPS 2023 · 6 citations
- Optimal Non-parametric Learning in Repeated Contextual Auctions with Strategic BuyerAlexey DrutsaICML 2020 · 18 citations
- Learning to Price Against a Moving TargetRenato Paes Leme, Balasubramanian Sivan, Yifeng Teng, Pratik WorahICML 2021 · 8 citations
