Locally Regularized Neural Differential Equations: Some Black Boxes were meant to remain closed!
Avik Pal, Alan Edelman, Christopher Vincent Rackauckas
Abstract
Implicit layer deep learning techniques, like Neural Differential Equations, have become an important modeling framework due to their ability to adapt to new problems automatically. Training a neural differential equation is effectively a search over a space of plausible dynamical systems. However, controlling the computational cost for these models is difficult since it relies on the number of steps the adaptive solver takes. Most prior works have used higher-order methods to reduce prediction timings while greatly increasing training time or reducing both training and prediction timings by relying on specific training algorithms, which are harder to use as a drop-in replacement due to strict requirements on automatic differentiation. In this manuscript, we use internal cost heuristics of adaptive differential equation solvers at stochastic time-points to guide the training towards learning a dynamical system that is easier to integrate. We "close the blackbox" and allow the use of our method with any adjoint technique for gradient calculations of the differential equation solution. We perform experimental studies to compare our method to global regularization to show that we attain similar performance numbers without compromising on the flexibility of implementation on ordinary differential equations (ODEs) and stochastic differential equations (SDEs). We develop two sampling strategies to trade-off between performance and training time. Our method reduces the number of function evaluations to 0 .556 × -0 .733 × and accelerates predictions by 1 .3 × -2 ×.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6419faf6-72c8-4a9d-ba19-466bb2ddf659Cited by top-tier papers1
Ask how each one uses itBuilds on4
- Multiscale Deep Equilibrium ModelsShaojie Bai, Vladlen Koltun, J. Zico KolterNeurIPS 2020 · 272 citations
- Learning Differential Equations that are Easy to SolveJacob Kelly, Jesse Bettencourt, Matthew J. Johnson, David DuvenaudNeurIPS 2020 · 134 citations
- Heavy Ball Neural Ordinary Differential EquationsHedi Xia, Vai Suliafu, Hangjie Ji, Tan M. Nguyen et al.NeurIPS 2021 · 75 citations
- Opening the Blackbox: Accelerating Neural Differential Equations by Regularizing Internal Solver HeuristicsAvik Pal, Yingbo Ma, Viral B. Shah, Christopher Vincent RackauckasICML 2021 · 44 citations
Related papers
- STEER : Simple Temporal Regularization For Neural ODEArnab Ghosh, Harkirat S. Behl, Emilien Dupont, Philip H. S. Torr et al.NeurIPS 2020 · 88 citations
- "Hey, that's not an ODE": Faster ODE Adjoints via SeminormsPatrick Kidger, Ricky T. Q. Chen, Terry J. LyonsICML 2021 · 56 citations
- Second-Order Neural ODE OptimizerGuan-Horng Liu, Tianrong Chen, Evangelos A. TheodorouNeurIPS 2021 · 20 citations
- Do Residual Neural Networks discretize Neural Ordinary Differential Equations?Michael E. Sander, Pierre Ablin, Gabriel PeyréNeurIPS 2022 · 42 citations
- How to Train Your Neural ODE: the World of Jacobian and Kinetic RegularizationChris Finlay, Jörn-Henrik Jacobsen, Levon Nurbekyan, Adam M. ObermanICML 2020 · 76 citations
