Error Propagation in Dynamic Programming: From Stochastic Control to American Option Pricing
Andrea Della Vecchia, Damir Filipovic
摘要
This paper investigates theoretical and methodological foundations for stochastic optimal control (SOC) in discrete time. We start formulating the control problem in a general dynamic programming framework, introducing the mathematical structure needed for a detailed convergence analysis. The associate value function is estimated through a sequence of approximations combining nonparametric regression methods and Monte Carlo subsampling. The regression step is performed within reproducing kernel Hilbert spaces (RKHSs), exploiting the classical KRR algorithm, while Monte Carlo sampling methods are introduced to estimate the continuation value. To assess the accuracy of our value function estimator, we propose a natural error decomposition and rigorously control the resulting error terms at each time step. We then analyze how this error propagates backward in time-from maturity to the initial stage-a relatively underexplored aspect of the SOC literature. Finally, we illustrate how our analysis naturally applies to a key financial application: the pricing of American options.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper5
- Path Integral Sampler: A Stochastic Control Approach For SamplingQinsheng Zhang, Yongxin ChenICLR 2022 · 被引用 177 次
- Kernel Methods Through the Roof: Handling Billions of Points EfficientlyGiacomo Meanti, Luigi Carratino, Lorenzo Rosasco, Alessandro RudiNeurIPS 2020 · 被引用 138 次
- Improved sampling via learned diffusionsLorenz Richter, Julius BernerICLR 2024 · 被引用 103 次
- Stochastic Optimal Control for Collective Variable Free Sampling of Molecular Transition PathsLars Holdijk, Yuanqi Du, Ferry Hooft, Priyank Jaini 等NeurIPS 2023 · 被引用 55 次
- Stochastic Optimal Control MatchingCarles Domingo-Enrich, Jiequn Han, Brandon Amos, Joan Bruna 等NeurIPS 2024 · 被引用 49 次
相关 Paper
- A Non-asymptotic Analysis of Non-parametric Temporal-Difference LearningEloïse Berthier, Ziad Kobeissi, Francis R. BachNeurIPS 2022 · 被引用 6 次
- Information Theoretic Regret Bounds for Online Nonlinear ControlSham M. Kakade, Akshay Krishnamurthy, Kendall Lowrey, Motoya Ohnishi 等NeurIPS 2020 · 被引用 137 次
- Policy Newton Algorithm in Reproducing Kernel Hilbert SpaceYixian Zhang, Huaze Tang, Changxu Wei, Chao Wang 等ICLR 2026 · 被引用 3 次
- Impact of Computation in Integral Reinforcement Learning for Continuous-Time ControlWenhan Cao, Wei PanICLR 2024 · 被引用 1 次
- Koopman Kernel RegressionPetar Bevanda, Max Beier, Armin Lederer, Stefan Sosnowski 等NeurIPS 2023 · 被引用 36 次
