Near-optimal Sketchy Natural Gradients for Physics-Informed Neural Networks
Maricela Best McKay, Avleen Kaur, Chen Greif, Brian Wetton
Abstract
Natural gradient methods for PINNs have achieved state-of-the-art performance with errors several orders of magnitude smaller than those achieved by standard optimizers such as ADAM or L-BFGS. However, computing natural gradients for PINNs is prohibitively computationally costly and memory-intensive for all but small neural network architectures. We develop a randomized algorithm for natural gradient descent for PINNs that uses sketching to approximate the natural gradient descent direction. We prove that the change of coordinate Gram matrix used in a natural gradient descent update has rapidly-decaying eigenvalues for a one-layer, one-dimensional neural network and empirically demonstrate that this structure holds for four different example problems. Under this structure, our sketching algorithm is guaranteed to provide a near-optimal lowrank approximation of the Gramian. Our algorithm dramatically speeds up computation time and reduces memory overhead. Additionally, in our experiments, the sketched natural gradient outperforms the original natural gradient in terms of accuracy, often achieving an error that is an order of magnitude smaller. Training time for a network with around 5,000 parameters is reduced from several hours to under two minutes. Training can be practically scaled to large network sizes; we optimize a PINN for a network with over a million parameters within a few minutes, a task for which the full Gram matrix does not fit in memory.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3dd6bad8-8bf2-4457-a2c8-dec5eaece15eCited by top-tier papers1
Ask how each one uses itBuilds on7
- Characterizing possible failure modes in physics-informed neural networksAditi S. Krishnapriyan, Amir Gholami, Shandian Zhe, Robert M. Kirby et al.NeurIPS 2021 · 1,421 citations
- Challenges in Training PINNs: A Loss Landscape PerspectivePratik Rathore, Weimu Lei, Zachary Frangella, Lu Lu et al.ICML 2024 · 137 citations
- Mitigating Propagation Failures in Physics-informed Neural Networks using Retain-Resample-Release (R3) SamplingArka Daw, Jie Bu, Sifan Wang, Paris Perdikaris et al.ICML 2023 · 95 citations
- Spectral Bias in Practice: The Role of Function Frequency in GeneralizationSara Fridovich-Keil, Raphael Gontijo Lopes, Rebecca RoelofsNeurIPS 2022 · 61 citations
- Provable Benefits of Overparameterization in Model Compression: From Double Descent to Pruning Neural NetworksXiangyu Chang, Yingcong Li, Samet Oymak, Christos ThrampoulidisAAAI 2021 · 58 citations
Related papers
- Fast Convergence of Natural Gradient Descent for Over-parameterized Physics-Informed Neural NetworksXianliang Xu, Wang Kong, Jiaheng Mao, Zhongyi Huang et al.ICLR 2026 · 6 citations
- Achieving High Accuracy with PINNs via Energy Natural Gradient DescentJohannes Müller, Marius ZeinhoferICML 2023 · 13 citations
- Implicit Stochastic Gradient Descent for Training Physics-Informed Neural NetworksYe Li, Songcan Chen, Sheng-Jun HuangAAAI 2023 · 5 citations
- ANaGRAM: A Natural Gradient Relative to Adapted Model for efficient PINNs learningNilo Schwencke, Cyril FurtlehnerICLR 2025
- A Layer-Wise Natural Gradient Optimizer for Training Deep Neural NetworksXiaolei Liu, Shaoshuai Li, Kaixin Gao, Binfeng WangNeurIPS 2024 · 2 citations
