Data-driven Optimal Filtering for Linear Systems with Unknown Noise Covariances
Shahriar Talebi, Amirhossein Taghvaei, Mehran Mesbahi
Abstract
This paper examines learning the optimal filtering policy, known as the Kalman gain, for a linear system with unknown noise covariance matrices using noisy output data. The learning problem is formulated as a stochastic policy optimization problem, aiming to minimize the output prediction error. This formulation provides a direct bridge between data-driven optimal control and, its dual, optimal filtering. Our contributions are twofold. Firstly, we conduct a thorough convergence analysis of the stochastic gradient descent algorithm, adopted for the filtering problem, accounting for biased gradients and stability constraints. Secondly, we carefully leverage a combination of tools from linear system theory and high-dimensional statistics to derive bias-variance error bounds that scale logarithmically with problem dimension, and, in contrast to subspace methods, the length of output trajectories only affects the bias term.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 51677cf9-259c-4e84-8353-b8aaf173d44dCited by top-tier papers1
Ask how each one uses itBuilds on2
- Logarithmic Regret Bound in Partially Observable Linear Dynamical SystemsSahin Lale, Kamyar Azizzadenesheli, Babak Hassibi, Anima AnandkumarNeurIPS 2020 · 106 citations
- Globally Convergent Policy Search for Output EstimationJack Umenberger, Max Simchowitz, Juan C. Perdomo, Kaiqing Zhang et al.NeurIPS 2022 · 16 citations
Related papers
- A New Approach to Learning Linear Dynamical SystemsAinesh Bakshi, Allen Liu, Ankur Moitra, Morris YauSTOC 2023 · 10 citations
- Online Policy Gradient for Model Free Learning of Linear Quadratic Regulators with √T RegretAsaf B. Cassel, Tomer KorenICML 2021 · 20 citations
- Stochastic Optimal Control and Estimation with Multiplicative and Internal NoiseFrancesco Damiani, Akiyuki Anzai, Jan Drugowitsch, Gregory C. DeAngelis et al.NeurIPS 2024 · 2 citations
- Logarithmic Regret for Learning Linear Quadratic Regulators EfficientlyAsaf B. Cassel, Alon Cohen, Tomer KorenICML 2020 · 68 citations
- Implicit Maximum a Posteriori Filtering via Adaptive OptimizationGianluca M. Bencomo, Jake Snell, Thomas L. GriffithsICLR 2024 · 4 citations
