Generalized Teacher Forcing for Learning Chaotic Dynamics
Florian Hess, Zahra Monfared, Manuel Brenner, Daniel Durstewitz
Abstract
Chaotic dynamical systems (DS) are ubiquitous in nature and society. Often we are interested in reconstructing such systems from observed time series for prediction or mechanistic insight, where by reconstruction we mean learning geometrical and invariant temporal properties of the system in question (like attractors). However, training reconstruction algorithms like recurrent neural networks (RNNs) on such systems by gradient-descent based techniques faces severe challenges. This is mainly due to exploding gradients caused by the exponential divergence of trajectories in chaotic systems. Moreover, for (scientific) interpretability we wish to have as low dimensional reconstructions as possible, preferably in a model which is mathematically tractable. Here we report that a surprisingly simple modification of teacher forcing leads to provably strictly all-time bounded gradients in training on chaotic systems, and, when paired with a simple architectural rearrangement of a tractable RNN design, piecewise-linear RNNs (PLRNNs), allows for faithful reconstruction in spaces of at most the dimensionality of the observed system. We show on several DS that with these amendments we can reconstruct DS better than current SOTA algorithms, in much lower dimensions. Performance differences were particularly compelling on real world data with which most other methods severely struggled. This work thus led to a simple yet powerful DS reconstruction algorithm which is highly interpretable at the same time.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e3c82187-0bfe-4238-bbcf-1c9f9e7b7c99Cited by top-tier papers26
- Training neural operators to preserve invariant measures of chaotic attractorsRuoxi Jiang, Peter Y. Lu, Elena Orlova, Rebecca WillettNeurIPS 2023 · 59 citations
- Inferring stochastic low-rank recurrent neural networks from neural dataMatthijs Pals, A Erdem Sagtekin, Felix Pei, Manuel Glöckler et al.NeurIPS 2024 · 37 citations
- Attractor Memory for Long-Term Time Series Forecasting: A Chaos PerspectiveJiaxi Hu, Yuehong Hu, Wei Chen, Ming Jin et al.NeurIPS 2024 · 35 citations
- Out-of-Domain Generalization in Dynamical Systems ReconstructionNiclas Alexander Göring, Florian Hess, Manuel Brenner, Zahra Monfared et al.ICML 2024 · 31 citations
- DySLIM: Dynamics Stable Learning by Invariant Measure for Chaotic SystemsYair Schiff, Zhong Yi Wan, Jeffrey B. Parker, Stephan Hoyer et al.ICML 2024 · 30 citations
Builds on13
- Fourier Neural Operator for Parametric Partial Differential EquationsZongyi Li, Nikola Borislavov Kovachki, Kamyar Azizzadenesheli, Burigede Liu et al.ICLR 2021 · 3,911 citations
- On the Variance of the Adaptive Learning Rate and BeyondLiyuan Liu, Haoming Jiang, Pengcheng He, Weizhu Chen et al.ICLR 2020 · 2,210 citations
- Coupled Oscillatory Recurrent Neural Network (coRNN): An accurate and (gradient) stable architecture for learning long time dependenciesT. Konstantin Rusch, Siddhartha MishraICLR 2021 · 121 citations
- On the difficulty of learning chaotic dynamics with RNNsJonas M. Mikhaeil, Zahra Monfared, Daniel DurstewitzNeurIPS 2022 · 109 citations
- Long Expressive Memory for Sequence ModelingT. Konstantin Rusch, Siddhartha Mishra, N. Benjamin Erichson, Michael W. MahoneyICLR 2022 · 57 citations
Related papers
- Tractable Dendritic RNNs for Reconstructing Nonlinear Dynamical SystemsManuel Brenner, Florian Hess, Jonas M. Mikhaeil, Leonard F. Bereska et al.ICML 2022 · 48 citations
- Almost-Linear RNNs Yield Highly Interpretable Symbolic Codes in Dynamical Systems ReconstructionManuel Brenner, Christoph Jürgen Hemmer, Zahra Monfared, Daniel DurstewitzNeurIPS 2024 · 22 citations
- Identifying nonlinear dynamical systems with multiple time scales and long-range dependenciesDominik Schmidt, Georgia Koppe, Zahra Monfared, Max Beutelspacher et al.ICLR 2021 · 41 citations
- Continuous-Time Piecewise-Linear Recurrent Neural NetworksAlena Brändle, Lukas Eisenmann, Florian Götz, Daniel DurstewitzICML 2026 · 2 citations
- Detecting Invariant Manifolds in ReLU-Based RNNsLukas Eisenmann, Alena Brändle, Zahra Monfared, Daniel DurstewitzICLR 2026 · 3 citations
