Randomized Sparse Neural Galerkin Schemes for Solving Evolution Equations with Deep Networks
Jules Berman, Benjamin Peherstorfer
摘要
Training neural networks sequentially in time to approximate solution fields of timedependent partial differential equations can be beneficial for preserving causality and other physics properties; however, the sequential-in-time training is numerically challenging because training errors quickly accumulate and amplify over time. This work introduces Neural Galerkin schemes that update randomized sparse subsets of network parameters at each time step. The randomization avoids overfitting locally in time and so helps prevent the error from accumulating quickly over the sequential-in-time training, which is motivated by dropout that addresses a similar issue of overfitting due to neuron co-adaptation. The sparsity of the update reduces the computational costs of training without losing expressiveness because many of the network parameters are redundant locally at each time step. In numerical experiments with a wide range of evolution equations, the proposed scheme with randomized sparse updates is up to two orders of magnitude more accurate at a fixed computational budget and up to two orders of magnitude faster at a fixed accuracy than schemes with dense updates. Preprint. Under review.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- CoLoRA: Continuous low-rank adaptation for reduced implicit neural modeling of parameterized partial differential equationsJules Berman, Benjamin PeherstorferICML 2024 · 被引用 17 次
- Fast training of accurate physics-informed neural networks without gradient descentChinmay Datar, Taniya Kapoor, Abhishek Chandra, Qing Sun 等ICLR 2026 · 被引用 10 次
- A Dirac-Frenkel-Onsager principle: Instantaneous residual minimization with gauge momentum for nonlinear parametrizations of PDE solutionsMatteo Raviola, Benjamin PeherstorferICML 2026 · 被引用 1 次
- TENG: Time-Evolving Natural Gradient for Solving PDEs With Deep Neural Nets Toward Machine PrecisionZhuo Chen, Jacob McCarran, Esteban Vizcaino, Marin Soljacic 等ICML 2024
它引用的顶会 Paper6
- Fourier Neural Operator for Parametric Partial Differential EquationsZongyi Li, Nikola Borislavov Kovachki, Kamyar Azizzadenesheli, Burigede Liu 等ICLR 2021 · 被引用 3,911 次
- Characterizing possible failure modes in physics-informed neural networksAditi S. Krishnapriyan, Amir Gholami, Shandian Zhe, Robert M. Kirby 等NeurIPS 2021 · 被引用 1,421 次
- Training Neural Networks with Fixed Sparse MasksYi-Lin Sung, Varun Nair, Colin RaffelNeurIPS 2021 · 被引用 295 次
- Mixout: Effective Regularization to Finetune Large-scale Pretrained Language ModelsCheolhyoung Lee, Kyunghyun Cho, Wanmo KangICLR 2020 · 被引用 233 次
- Rational neural networksNicolas Boullé, Yuji Nakatsukasa, Alex TownsendNeurIPS 2020 · 被引用 130 次
相关 Paper
- SparseProp: Efficient Event-Based Simulation and Training of Sparse Recurrent Spiking Neural NetworksRainer EngelkenNeurIPS 2023 · 被引用 15 次
- Neural Stochastic PDEs: Resolution-Invariant Learning of Continuous Spatiotemporal DynamicsCristopher Salvi, Maud Lemercier, Andris GerasimovicsNeurIPS 2022 · 被引用 70 次
- RSC: Accelerate Graph Neural Networks Training via Randomized Sparse ComputationsZirui Liu, Shengyuan Chen, Kaixiong Zhou, Daochen Zha 等ICML 2023 · 被引用 24 次
- Neural Spectral Methods: Self-supervised learning in the spectral domainYiheng Du, Nithin Chalapathi, Aditi S. KrishnapriyanICLR 2024 · 被引用 16 次
- Efficient Neural Network Training via Forward and Backward Propagation SparsificationXiao Zhou, Weizhong Zhang, Zonghao Chen, Shizhe Diao 等NeurIPS 2021 · 被引用 57 次
