When Demands Evolve Larger and Noisier: Learning and Earning in a Growing Environment
Feng Zhu, Zeyu Zheng
摘要
We consider a single-product dynamic pricing problem under a specific non-stationary setting, where the underlying demand process grows over time in expectation and also possibly in the level of random fluctuation. The decision maker sequentially sets price in each time period and learns the unknown demand model, with the goal of maximizing expected cumulative revenue over a time horizon T . We prove matching upper and lower bounds on regret and provide near-optimal pricing policies. We show how the growth rate of random fluctuation over time affects the best achievable regret order and the near-optimal policy design. In the analysis, we show that whether the seller knows the length of time horizon T in advance or not surprisingly render different optimal regret orders. We then extend the demand model such that the optimal price may vary with time and present a novel and near-optimal policy for the extended model. Finally, we consider an analogous nonstationary setting in the canonical multi-armed bandit problem, and points out that knowing or not knowing the length of time horizon T render the same optimal regret order, in contrast to the non-stationary dynamic pricing problem.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Context-Based Dynamic Pricing with Partially Linear Demand ModelJinzhi Bu, David Simchi-Levi, Chonghuan WangNeurIPS 2022 · 被引用 19 次
- Non-stationary Experimental Design under Linear TrendsDavid Simchi-Levi, Chonghuan Wang, Zeyu ZhengNeurIPS 2023 · 被引用 6 次
- Dynamic Planning and Learning under Recovering RewardsDavid Simchi-Levi, Zeyu Zheng, Feng ZhuICML 2021 · 被引用 6 次
- Bandits with Knapsacks: Advice on Time-Varying DemandsLixing Lyu, Wang Chi CheungICML 2023 · 被引用 4 次
- Dynamic Service Fee Pricing under Strategic Behavior: Actions as Instruments and Phase TransitionRui Ai, David Simchi-Levi, Feng ZhuNeurIPS 2024
它引用的顶会 Paper1
相关 Paper
- Dynamic pricing and assortment under a contextual MNL demandNoémie Périvier, Vineet GoyalNeurIPS 2022 · 被引用 29 次
- Learning to Price Against a Moving TargetRenato Paes Leme, Balasubramanian Sivan, Yifeng Teng, Pratik WorahICML 2021 · 被引用 8 次
- Online Second Price Auction with Semi-Bandit Feedback under the Non-Stationary SettingHaoyu Zhao, Wei ChenAAAI 2020 · 被引用 15 次
- Tightening Regret Lower and Upper Bounds in Restless Rising BanditsCristiano Migali, Marco Mussi, Gianmarco Genalti, Alberto Maria MetelliNeurIPS 2025
- Improved Algorithms for Contextual Dynamic PricingMatilde Tullii, Solenne Gaucher, Nadav Merlis, Vianney PerchetNeurIPS 2024 · 被引用 18 次
