How to Train Your Neural ODE: the World of Jacobian and Kinetic Regularization
Chris Finlay, Jörn-Henrik Jacobsen, Levon Nurbekyan, Adam M. Oberman
Abstract
Training neural ODEs on large datasets has not been tractable due to the necessity of allowing the adaptive numerical ODE solver to refine its step size to very small values. In practice this leads to dynamics equivalent to many hundreds or even thousands of layers. In this paper, we overcome this apparent difficulty by introducing a theoretically-grounded combination of both optimal transport and stability regularizations which encourage neural ODEs to prefer simpler dynamics out of all the dynamics that solve a problem well. Simpler dynamics lead to faster convergence and to fewer discretizations of the solver, considerably decreasing wall-clock time without loss in performance. Our approach allows us to train neural ODE-based generative models to the same performance as the unregularized dynamics, with significant reductions in training time. This brings neural ODEs closer to practical relevance in large-scale applications.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2576d226-d59b-45c5-87fa-b16e077770faCited by top-tier papers94
- HiPPO: Recurrent Memory with Optimal Polynomial ProjectionsAlbert Gu, Tri Dao, Stefano Ermon, Atri Rudra et al.NeurIPS 2020 · 1,100 citations
- GRAND: Graph Neural DiffusionBen Chamberlain, James Rowbottom, Maria I. Gorinova, Michael M. Bronstein et al.ICML 2021 · 358 citations
- E(n) Equivariant Normalizing FlowsVictor Garcia Satorras, Emiel Hoogeboom, Fabian Fuchs, Ingmar Posner et al.NeurIPS 2021 · 246 citations
- Multisample Flow Matching: Straightening Flows with Minibatch CouplingsAram-Alexandre Pooladian, Heli Ben-Hamu, Carles Domingo-Enrich, Brandon Amos et al.ICML 2023 · 243 citations
- OT-Flow: Fast and Accurate Continuous Normalizing Flows via Optimal TransportDerek Onken, Samy Wu Fung, Xingjian Li, Lars RuthottoAAAI 2021 · 210 citations
Related papers
- TO-FLOW: Efficient Continuous Normalizing Flows with Temporal Optimization adjoint with Moving SpeedShian Du, Yihong Luo, Wei Chen, Jian Xu et al.CVPR 2022 · 2 citations
- STEER : Simple Temporal Regularization For Neural ODEArnab Ghosh, Harkirat S. Behl, Emilien Dupont, Philip H. S. Torr et al.NeurIPS 2020 · 88 citations
- Stability-Informed Initialization of Neural Ordinary Differential EquationsTheodor Westny, Arman Mohammadi, Daniel Jung, Erik FriskICML 2024 · 6 citations
- Training Generative Adversarial Networks by Solving Ordinary Differential EquationsChongli Qin, Yan Wu, Jost Tobias Springenberg, Andy Brock et al.NeurIPS 2020 · 35 citations
- Opening the Blackbox: Accelerating Neural Differential Equations by Regularizing Internal Solver HeuristicsAvik Pal, Yingbo Ma, Viral B. Shah, Christopher Vincent RackauckasICML 2021 · 44 citations
