Randomized Sparse Neural Galerkin Schemes for Solving Evolution Equations with Deep Networks
Jules Berman, Benjamin Peherstorfer
Abstract
Training neural networks sequentially in time to approximate solution fields of timedependent partial differential equations can be beneficial for preserving causality and other physics properties; however, the sequential-in-time training is numerically challenging because training errors quickly accumulate and amplify over time. This work introduces Neural Galerkin schemes that update randomized sparse subsets of network parameters at each time step. The randomization avoids overfitting locally in time and so helps prevent the error from accumulating quickly over the sequential-in-time training, which is motivated by dropout that addresses a similar issue of overfitting due to neuron co-adaptation. The sparsity of the update reduces the computational costs of training without losing expressiveness because many of the network parameters are redundant locally at each time step. In numerical experiments with a wide range of evolution equations, the proposed scheme with randomized sparse updates is up to two orders of magnitude more accurate at a fixed computational budget and up to two orders of magnitude faster at a fixed accuracy than schemes with dense updates. Preprint. Under review.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers4
- CoLoRA: Continuous low-rank adaptation for reduced implicit neural modeling of parameterized partial differential equationsJules Berman, Benjamin PeherstorferICML 2024 · 17 citations
- Fast training of accurate physics-informed neural networks without gradient descentChinmay Datar, Taniya Kapoor, Abhishek Chandra, Qing Sun et al.ICLR 2026 · 10 citations
- A Dirac-Frenkel-Onsager principle: Instantaneous residual minimization with gauge momentum for nonlinear parametrizations of PDE solutionsMatteo Raviola, Benjamin PeherstorferICML 2026 · 1 citation
- TENG: Time-Evolving Natural Gradient for Solving PDEs With Deep Neural Nets Toward Machine PrecisionZhuo Chen, Jacob McCarran, Esteban Vizcaino, Marin Soljacic et al.ICML 2024
Builds on6
- Fourier Neural Operator for Parametric Partial Differential EquationsZongyi Li, Nikola Borislavov Kovachki, Kamyar Azizzadenesheli, Burigede Liu et al.ICLR 2021 · 3,911 citations
- Characterizing possible failure modes in physics-informed neural networksAditi S. Krishnapriyan, Amir Gholami, Shandian Zhe, Robert M. Kirby et al.NeurIPS 2021 · 1,421 citations
- Training Neural Networks with Fixed Sparse MasksYi-Lin Sung, Varun Nair, Colin RaffelNeurIPS 2021 · 295 citations
- Mixout: Effective Regularization to Finetune Large-scale Pretrained Language ModelsCheolhyoung Lee, Kyunghyun Cho, Wanmo KangICLR 2020 · 233 citations
- Rational neural networksNicolas Boullé, Yuji Nakatsukasa, Alex TownsendNeurIPS 2020 · 130 citations
Related papers
- SparseProp: Efficient Event-Based Simulation and Training of Sparse Recurrent Spiking Neural NetworksRainer EngelkenNeurIPS 2023 · 15 citations
- Neural Stochastic PDEs: Resolution-Invariant Learning of Continuous Spatiotemporal DynamicsCristopher Salvi, Maud Lemercier, Andris GerasimovicsNeurIPS 2022 · 70 citations
- RSC: Accelerate Graph Neural Networks Training via Randomized Sparse ComputationsZirui Liu, Shengyuan Chen, Kaixiong Zhou, Daochen Zha et al.ICML 2023 · 24 citations
- Neural Spectral Methods: Self-supervised learning in the spectral domainYiheng Du, Nithin Chalapathi, Aditi S. KrishnapriyanICLR 2024 · 16 citations
- Efficient Neural Network Training via Forward and Backward Propagation SparsificationXiao Zhou, Weizhong Zhang, Zonghao Chen, Shizhe Diao et al.NeurIPS 2021 · 57 citations
