Optimizing Neural Networks via Koopman Operator Theory
Akshunna S. Dogra, William T. Redman
摘要
Koopman operator theory, a powerful framework for discovering the underlying dynamics of nonlinear dynamical systems, was recently shown to be intimately connected with neural network training. In this work, we take the first steps in making use of this connection. As Koopman operator theory is a linear theory, a successful implementation of it in evolving network weights and biases offers the promise of accelerated training, especially in the context of deep networks, where optimization is inherently a non-convex problem. We show that Koopman operator theoretic methods allow for accurate predictions of weights and biases of feedforward, fully connected deep networks over a non-trivial range of training time. During this window, we find that our approach is >10x faster than various gradient descent based methods (e.g. Adam, Adadelta, Adagrad), in line with our complexity analysis. We end by highlighting open questions in this exciting intersection between dynamical systems and neural network theory. We highlight additional methods by which our results could be expanded to broader classes of networks and larger training intervals, which shall be the focus of future work.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper16
- Noisy Recurrent Neural NetworksSoon Hoe Lim, N. Benjamin Erichson, Liam Hodgkinson, Michael W. MahoneyNeurIPS 2021 · 被引用 77 次
- Generative Modeling of Regular and Irregular Time Series Data via Koopman VAEsIlan Naiman, N. Benjamin Erichson, Pu Ren, Michael W. Mahoney 等ICLR 2024 · 被引用 49 次
- Learning Physics Constrained Dynamics Using AutoencodersTsung-Yen Yang, Justinian Rosca, Karthik Narasimhan, Peter J. RamadgeNeurIPS 2022 · 被引用 39 次
- Memory-Efficient Learning of Stable Linear Dynamical Systems for Prediction and ControlGiorgos Mamakoukas, Orest Xherija, Todd D. MurpheyNeurIPS 2020 · 被引用 25 次
- An Operator Theoretic View On Pruning Deep Neural NetworksWilliam T. Redman, Maria Fonoberova, Ryan Mohr, Yannis G. Kevrekidis 等ICLR 2022 · 被引用 21 次
相关 Paper
- Predictive Differential Training Guided by Training DynamicsFanqi Wang, Weisheng Tang, Landon Harris, Hairong Qi 等ICLR 2026
- Identifying Equivalent Training DynamicsWilliam T. Redman, Juan M. Bello-Rivas, Maria Fonoberova, Ryan Mohr 等NeurIPS 2024 · 被引用 15 次
- Koopman-based generalization bound: New aspect for full-rank weightsYuka Hashimoto, Sho Sonoda, Isao Ishikawa, Atsushi Nitanda 等ICLR 2024 · 被引用 6 次
- An Operator Theoretic Approach for Analyzing Sequence Neural NetworksIlan Naiman, Omri AzencotAAAI 2023 · 被引用 14 次
- SKOLR: Structured Koopman Operator Linear RNN for Time-Series ForecastingYitian Zhang, Liheng Ma, Antonios Valkanas, Boris N. Oreshkin 等ICML 2025
