Long-Context Linear System Identification
Oguz Kaan Yüksel, Mathieu Even, Nicolas Flammarion
Abstract
This paper addresses the problem of long-context linear system identification, where the state x t of a dynamical system at time t depends linearly on previous states x s over a fixed context window of length p. We establish a sample complexity bound that matches the i.i.d. parametric rate up to logarithmic factors for a broad class of systems, extending previous works that considered only first-order dependencies. Our findings reveal a "learning-without-mixing" phenomenon, indicating that learning long-context linear autoregressive models is not hindered by slow mixing properties potentially associated with extended context windows. Additionally, we extend these results to (i) shared low-rank representations, where rank-regularized estimators improve the dependence of the rates on the dimensionality, and (ii) misspecified context lengths in strictly stable systems, where shorter contexts offer statistical advantages.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 51ebd32c-2ec9-42bb-907d-4d24262de374Cited by top-tier papers2
- In-context Learning of Linear Dynamical Systems with Transformers: Approximation Bounds and Depth-separationFrank Cole, Yuxuan Zhao, Yulong Lu, Tianhao ZhangNeurIPS 2025 · 1 citation
- Induction Heads Interpolate N-GramsFrancesco D'Angelo, Oğuz Yüksel, Swathi Narashiman, Nicolas FlammarionICML 2026
Builds on9
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Provable Meta-Learning of Linear RepresentationsNilesh Tripuraneni, Chi Jin, Michael I. JordanICML 2021 · 218 citations
- Scalable Pre-training of Large Autoregressive Image ModelsAlaaeldin El-Nouby, Michal Klein, Shuangfei Zhai, Miguel Ángel Bautista et al.ICML 2024 · 130 citations
- Learning Dynamical Systems via Koopman Operator Regression in Reproducing Kernel Hilbert SpacesVladimir Kostic, Pietro Novelli, Andreas Maurer, Carlo Ciliberto et al.NeurIPS 2022 · 109 citations
- Learning with little mixingIngvar M. Ziemann, Stephen TuNeurIPS 2022 · 41 citations
Related papers
- Stochastic Contextual Bandits with Long Horizon RewardsYuzhen Qin, Yingcong Li, Fabio Pasqualetti, Maryam Fazel et al.AAAI 2023 · 3 citations
- A New Approach to Learning Linear Dynamical SystemsAinesh Bakshi, Allen Liu, Ankur Moitra, Morris YauSTOC 2023 · 10 citations
- On learning linear dynamical systems in context with attention layersMaria-Luiza Vladarean, Xuhui Zhang, Suvrit SraICLR 2026
- Learning Low-dimensional Latent Dynamics from High-dimensional Observations: Non-asymptotics and Lower BoundsYuyang Zhang, Shahriar Talebi, Na LiICML 2024 · 5 citations
- PAC-Bayesian Error Bound, via Rényi Divergence, for a Class of Linear Time-Invariant State-Space ModelsDeividas Eringis, John Leth, Zheng-Hua Tan, Rafal Wisniewski et al.ICML 2024 · 2 citations
