Optimizing Neural Networks via Koopman Operator Theory
Akshunna S. Dogra, William T. Redman
Abstract
Koopman operator theory, a powerful framework for discovering the underlying dynamics of nonlinear dynamical systems, was recently shown to be intimately connected with neural network training. In this work, we take the first steps in making use of this connection. As Koopman operator theory is a linear theory, a successful implementation of it in evolving network weights and biases offers the promise of accelerated training, especially in the context of deep networks, where optimization is inherently a non-convex problem. We show that Koopman operator theoretic methods allow for accurate predictions of weights and biases of feedforward, fully connected deep networks over a non-trivial range of training time. During this window, we find that our approach is >10x faster than various gradient descent based methods (e.g. Adam, Adadelta, Adagrad), in line with our complexity analysis. We end by highlighting open questions in this exciting intersection between dynamical systems and neural network theory. We highlight additional methods by which our results could be expanded to broader classes of networks and larger training intervals, which shall be the focus of future work.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers16
- Noisy Recurrent Neural NetworksSoon Hoe Lim, N. Benjamin Erichson, Liam Hodgkinson, Michael W. MahoneyNeurIPS 2021 · 77 citations
- Generative Modeling of Regular and Irregular Time Series Data via Koopman VAEsIlan Naiman, N. Benjamin Erichson, Pu Ren, Michael W. Mahoney et al.ICLR 2024 · 49 citations
- Learning Physics Constrained Dynamics Using AutoencodersTsung-Yen Yang, Justinian Rosca, Karthik Narasimhan, Peter J. RamadgeNeurIPS 2022 · 39 citations
- Memory-Efficient Learning of Stable Linear Dynamical Systems for Prediction and ControlGiorgos Mamakoukas, Orest Xherija, Todd D. MurpheyNeurIPS 2020 · 25 citations
- An Operator Theoretic View On Pruning Deep Neural NetworksWilliam T. Redman, Maria Fonoberova, Ryan Mohr, Yannis G. Kevrekidis et al.ICLR 2022 · 21 citations
Related papers
- Predictive Differential Training Guided by Training DynamicsFanqi Wang, Weisheng Tang, Landon Harris, Hairong Qi et al.ICLR 2026
- Identifying Equivalent Training DynamicsWilliam T. Redman, Juan M. Bello-Rivas, Maria Fonoberova, Ryan Mohr et al.NeurIPS 2024 · 15 citations
- Koopman-based generalization bound: New aspect for full-rank weightsYuka Hashimoto, Sho Sonoda, Isao Ishikawa, Atsushi Nitanda et al.ICLR 2024 · 6 citations
- An Operator Theoretic Approach for Analyzing Sequence Neural NetworksIlan Naiman, Omri AzencotAAAI 2023 · 14 citations
- SKOLR: Structured Koopman Operator Linear RNN for Time-Series ForecastingYitian Zhang, Liheng Ma, Antonios Valkanas, Boris N. Oreshkin et al.ICML 2025
