Dynamic Pricing with Monotonicity Constraint under Unknown Parametric Demand Model
Su Jia, Andrew A. Li, R. Ravi
摘要
We consider a Continuum-Armed Bandit problem with an additional monotonicity constraint (or “markdown” constraint) on the actions selected. This problem faith-fully models a natural revenue management problem, called “markdown pricing”, where the objective is to adaptively reduce the price over a finite horizon to maximize the expected revenues. Chen ([3]) and Jia et al ([9]) recently showed a tight T 3 / 4 regret bound over T rounds under minimal assumptions of unimodality and Lipschitzness in the reward function. This bound shows that markdown pricing is strictly harder than unconstrained dynamic pricing (i.e., without the monotonicity constraint), which admits T 2 / 3 regret under the same assumptions ([11]). However, in practice, demand functions are usually assumed to have certain functional forms (e.g. linear or exponential), rendering the demand learning easier and suggesting better regret bounds. In this work we introduce a concept, markdown dimension , that measures the complexity of a parametric family, and present optimal regret bounds that improve upon the previous T 3 / 4 bound under this framework.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper1
相关 Paper
- When Demands Evolve Larger and Noisier: Learning and Earning in a Growing EnvironmentFeng Zhu, Zeyu ZhengICML 2020 · 被引用 15 次
- Contextual Dynamic Pricing with Unknown Noise: Explore-then-UCB Strategy and Improved RegretsYiyun Luo, Will Wei Sun, Yufeng LiuNeurIPS 2022 · 被引用 19 次
- Semi-Parametric Contextual Pricing with General SmoothnessYuxuan Han, Xiaocong Xu, Yuxiao Wen, Yanjun Han 等ICLR 2026
- Effective Dimension in Bandit Problems under CensorshipGauthier Guinet, Saurabh Amin, Patrick JailletNeurIPS 2022 · 被引用 3 次
- Improved Algorithms for Contextual Dynamic PricingMatilde Tullii, Solenne Gaucher, Nadav Merlis, Vianney PerchetNeurIPS 2024 · 被引用 18 次
