Sparse Flows: Pruning Continuous-depth Models
Lucas Liebenwein, Ramin M. Hasani, Alexander Amini, Daniela Rus
Abstract
Continuous deep learning architectures enable learning of flexible probabilistic models for predictive modeling as neural ordinary differential equations (ODEs), and for generative modeling as continuous normalizing flows. In this work, we design a framework to decipher the internal dynamics of these continuous depth models by pruning their network architectures. Our empirical results suggest that pruning improves generalization for neural ODEs in generative modeling. We empirically show that the improvement is because pruning helps avoid mode-collapse and flatten the loss surface. Moreover, pruning finds efficient neural ODE representations with up to 98% less parameters compared to the original network, without loss of accuracy. We hope our results will invigorate further research into the performance-size trade-offs of modern continuous-depth models.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3a337188-0d45-46bf-a951-51b70012d73bCited by top-tier papers6
- Causal Navigation by Continuous-time Neural NetworksCharles Vorbach, Ramin M. Hasani, Alexander Amini, Mathias Lechner et al.NeurIPS 2021 · 64 citations
- Compressing Neural Networks: Towards Determining the Optimal Layer-wise DecompositionLucas Liebenwein, Alaa Maalouf, Dan Feldman, Daniela RusNeurIPS 2021 · 60 citations
- Sparsity in Continuous-Depth Neural NetworksHananeh Aliee, Till Richter, Mikhail Solonin, Ignacio Ibarra et al.NeurIPS 2022 · 21 citations
- On the Forward Invariance of Neural ODEsWei Xiao, Tsun-Hsuan Wang, Ramin M. Hasani, Mathias Lechner et al.ICML 2023 · 17 citations
- Liquid Structural State-Space ModelsRamin M. Hasani, Mathias Lechner, Tsun-Hsuan Wang, Makram Chahine et al.ICLR 2023 · 13 citations
Builds on12
- Liquid Time-constant NetworksRamin M. Hasani, Mathias Lechner, Alexander Amini, Daniela Rus et al.AAAI 2021 · 399 citations
- Dissecting Neural ODEsStefano Massaroli, Michael Poli, Jinkyoo Park, Atsushi Yamashita et al.NeurIPS 2020 · 261 citations
- OT-Flow: Fast and Accurate Continuous Normalizing Flows via Optimal TransportDerek Onken, Samy Wu Fung, Xingjian Li, Lars RuthottoAAAI 2021 · 210 citations
- Provable Filter Pruning for Efficient Neural NetworksLucas Liebenwein, Cenk Baykal, Harry Lang, Dan Feldman et al.ICLR 2020 · 161 citations
- Coupling-based Invertible Neural Networks Are Universal Diffeomorphism ApproximatorsTakeshi Teshima, Isao Ishikawa, Koichi Tojo, Kenta Oono et al.NeurIPS 2020 · 129 citations
Related papers
- A shooting formulation of deep learningFrançois-Xavier Vialard, Roland Kwitt, Susan Wei, Marc NiethammerNeurIPS 2020 · 16 citations
- Characteristic Neural Ordinary Differential EquationXingzi Xu, Ali Hasan, Khalil Elkhalil, Jie Ding et al.ICLR 2023 · 2 citations
- Improving Neural ODE Training with Temporal Adaptive Batch NormalizationSu Zheng, Zhengqi Gao, Fan-Keng Sun, Duane S. Boning et al.NeurIPS 2024 · 5 citations
- Stateful ODE-Nets using Basis Function ExpansionsAlejandro F. Queiruga, N. Benjamin Erichson, Liam Hodgkinson, Michael W. MahoneyNeurIPS 2021 · 18 citations
- Symbolic Neural Ordinary Differential EquationsXin Li, Chengli Zhao, Xue Zhang, Xiaojun DuanAAAI 2025 · 3 citations
