Error Propagation in Dynamic Programming: From Stochastic Control to American Option Pricing
Andrea Della Vecchia, Damir Filipovic
Abstract
This paper investigates theoretical and methodological foundations for stochastic optimal control (SOC) in discrete time. We start formulating the control problem in a general dynamic programming framework, introducing the mathematical structure needed for a detailed convergence analysis. The associate value function is estimated through a sequence of approximations combining nonparametric regression methods and Monte Carlo subsampling. The regression step is performed within reproducing kernel Hilbert spaces (RKHSs), exploiting the classical KRR algorithm, while Monte Carlo sampling methods are introduced to estimate the continuation value. To assess the accuracy of our value function estimator, we propose a natural error decomposition and rigorously control the resulting error terms at each time step. We then analyze how this error propagates backward in time-from maturity to the initial stage-a relatively underexplored aspect of the SOC literature. Finally, we illustrate how our analysis naturally applies to a key financial application: the pricing of American options.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a7d372c7-1506-4d00-97bd-175ef566f580Builds on5
- Path Integral Sampler: A Stochastic Control Approach For SamplingQinsheng Zhang, Yongxin ChenICLR 2022 · 177 citations
- Kernel Methods Through the Roof: Handling Billions of Points EfficientlyGiacomo Meanti, Luigi Carratino, Lorenzo Rosasco, Alessandro RudiNeurIPS 2020 · 138 citations
- Improved sampling via learned diffusionsLorenz Richter, Julius BernerICLR 2024 · 103 citations
- Stochastic Optimal Control for Collective Variable Free Sampling of Molecular Transition PathsLars Holdijk, Yuanqi Du, Ferry Hooft, Priyank Jaini et al.NeurIPS 2023 · 55 citations
- Stochastic Optimal Control MatchingCarles Domingo-Enrich, Jiequn Han, Brandon Amos, Joan Bruna et al.NeurIPS 2024 · 49 citations
Related papers
- A Non-asymptotic Analysis of Non-parametric Temporal-Difference LearningEloïse Berthier, Ziad Kobeissi, Francis R. BachNeurIPS 2022 · 6 citations
- Information Theoretic Regret Bounds for Online Nonlinear ControlSham M. Kakade, Akshay Krishnamurthy, Kendall Lowrey, Motoya Ohnishi et al.NeurIPS 2020 · 137 citations
- Policy Newton Algorithm in Reproducing Kernel Hilbert SpaceYixian Zhang, Huaze Tang, Changxu Wei, Chao Wang et al.ICLR 2026 · 3 citations
- Impact of Computation in Integral Reinforcement Learning for Continuous-Time ControlWenhan Cao, Wei PanICLR 2024 · 1 citation
- Koopman Kernel RegressionPetar Bevanda, Max Beier, Armin Lederer, Stefan Sosnowski et al.NeurIPS 2023 · 36 citations
