Lune

NeurIPS2025Top-tier venue

Contextual Online Pricing with (Biased) Offline Data

Yixuan Zhang, Ruihao Zhu, Qiaomin Xie

2025Year
3Citations

Abstract

We study contextual online pricing with biased offline data. For the scalar price elasticity case, we identify the instance-dependent quantity δ2\delta^2 that measures how far the offline data lies from the (unknown) online optimum. We show that the time length TT, bias bound VV, size NN and dispersion λmin⁡(Σ^)\lambda_{\min}(\hat{\Sigma}) of the offline data, and δ2\delta^2 jointly determine the statistical complexity. An Optimism-in-the-Face-of-Uncertainty (OFU) policy achieves a minimax-optimal, instance-dependent regret bound O~(dT∧(V2T+dTλmin⁡(Σ^)+(N∧T)δ2))\tilde{\mathcal{O}}\big(d\sqrt{T} \wedge (V^2T + \frac{dT}{\lambda_{\min}(\hat{\Sigma}) + (N \wedge T) \delta^2})\big). For general price elasticity, we establish a worst-case, minimax-optimal rate O~(dT∧(V2T+dTλmin⁡(Σ^)))\tilde{\mathcal{O}}\big(d\sqrt{T} \wedge (V^2T + \frac{dT }{\lambda_{\min}(\hat{\Sigma})})\big) and provide a generalized OFU algorithm that attains it. When the bias bound VV is unknown, we design a robust variant that always guarantees sub-linear regret and strictly improves on purely online methods whenever the exact bias is small. These results deliver the first tight regret guarantees for contextual pricing in the presence of biased offline data. Our techniques also transfer verbatim to stochastic linear bandits with biased offline data, yielding analogous bounds.

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

lune papers fulltext 7a8bbf20-3184-4628-9f84-c9fd5db7f14e

Builds on5

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines