When Demands Evolve Larger and Noisier: Learning and Earning in a Growing Environment
Feng Zhu, Zeyu Zheng
Abstract
We consider a single-product dynamic pricing problem under a specific non-stationary setting, where the underlying demand process grows over time in expectation and also possibly in the level of random fluctuation. The decision maker sequentially sets price in each time period and learns the unknown demand model, with the goal of maximizing expected cumulative revenue over a time horizon T . We prove matching upper and lower bounds on regret and provide near-optimal pricing policies. We show how the growth rate of random fluctuation over time affects the best achievable regret order and the near-optimal policy design. In the analysis, we show that whether the seller knows the length of time horizon T in advance or not surprisingly render different optimal regret orders. We then extend the demand model such that the optimal price may vary with time and present a novel and near-optimal policy for the extended model. Finally, we consider an analogous nonstationary setting in the canonical multi-armed bandit problem, and points out that knowing or not knowing the length of time horizon T render the same optimal regret order, in contrast to the non-stationary dynamic pricing problem.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext fef8d813-8125-4f4a-ab6e-bc61dcc0c09cCited by top-tier papers5
- Context-Based Dynamic Pricing with Partially Linear Demand ModelJinzhi Bu, David Simchi-Levi, Chonghuan WangNeurIPS 2022 · 19 citations
- Non-stationary Experimental Design under Linear TrendsDavid Simchi-Levi, Chonghuan Wang, Zeyu ZhengNeurIPS 2023 · 6 citations
- Dynamic Planning and Learning under Recovering RewardsDavid Simchi-Levi, Zeyu Zheng, Feng ZhuICML 2021 · 6 citations
- Bandits with Knapsacks: Advice on Time-Varying DemandsLixing Lyu, Wang Chi CheungICML 2023 · 4 citations
- Dynamic Service Fee Pricing under Strategic Behavior: Actions as Instruments and Phase TransitionRui Ai, David Simchi-Levi, Feng ZhuNeurIPS 2024
Builds on1
Related papers
- Dynamic pricing and assortment under a contextual MNL demandNoémie Périvier, Vineet GoyalNeurIPS 2022 · 29 citations
- Learning to Price Against a Moving TargetRenato Paes Leme, Balasubramanian Sivan, Yifeng Teng, Pratik WorahICML 2021 · 8 citations
- Online Second Price Auction with Semi-Bandit Feedback under the Non-Stationary SettingHaoyu Zhao, Wei ChenAAAI 2020 · 15 citations
- Tightening Regret Lower and Upper Bounds in Restless Rising BanditsCristiano Migali, Marco Mussi, Gianmarco Genalti, Alberto Maria MetelliNeurIPS 2025
- Improved Algorithms for Contextual Dynamic PricingMatilde Tullii, Solenne Gaucher, Nadav Merlis, Vianney PerchetNeurIPS 2024 · 18 citations
