Efficient Continuous-Depth Modeling with GRU Equivalents
Ayan Banerjee, BIN XU, Sandeep Gupta
摘要
Continuous-Depth Neural Networks (CDNNs), including Neural Ordinary Differential Equations (ODEs) and Liquid-Time-Constant Neural Networks (LTC-NN), suffer from high computational costs due to solving numerous nonlinear ODEs during training and inference. We introduce Continuous Depth Acceleration (CoDA), a framework that leverages Mori–Zwanzig/Koopman operator theory to replace continuous-depth layers requiring multiple nonlinear ODEs with a compact GRU module, a single low-dimensional linear ODE, and a dense layer. We prove PAC learnability of CoDA, establishing that this transformation preserves accuracy and can be applied repeatedly across multiple layers with unified backpropagation. Experiments on the Liquid Foundation Model (LFM-1.2B) demonstrate training speedup and inference speedup without loss of accuracy. Across six real-world LTC-NN applications, CoDA consistently outperforms state-of-the-art acceleration techniques—including neural flows, model order reduction, and variational formulations—in both training and inference time while maintaining competitive or superior accuracy. The implementation and datasets are publicly available at https://github.com/ImpactLabASU/CoDA-ICML2026.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper5
- Liquid Time-constant NetworksRamin M. Hasani, Mathias Lechner, Alexander Amini, Daniela Rus 等AAAI 2021 · 被引用 399 次
- Deja Vu: Contextual Sparsity for Efficient LLMs at Inference TimeZichang Liu, Jue Wang, Tri Dao, Tianyi Zhou 等ICML 2023 · 被引用 318 次
- Neural Flows: Efficient Alternative to Neural ODEsMarin Bilos, Johanna Sommer, Syama Sundar Rangapuram, Tim Januschowski 等NeurIPS 2021 · 被引用 151 次
- WildChat-50M: A Deep Dive Into the Role of Synthetic Data in Post-TrainingBenjamin Feuer, Chinmay HegdeICML 2025
- Accelerating Neural ODEs: A Variational Formulation-based ApproachHongjue Zhao, Yuchen Wang, Hairong Qi, Zijie Huang 等ICLR 2025
相关 Paper
- Sparse Flows: Pruning Continuous-depth ModelsLucas Liebenwein, Ramin M. Hasani, Alexander Amini, Daniela RusNeurIPS 2021 · 被引用 21 次
- How Deep Do We Need: Accelerating Training and Inference of Neural ODEs via Control PerspectiveKeyan Miao, Konstantinos GatsisICML 2024 · 被引用 2 次
- Second-Order Neural ODE OptimizerGuan-Horng Liu, Tianrong Chen, Evangelos A. TheodorouNeurIPS 2021 · 被引用 20 次
- Improving Neural ODE Training with Temporal Adaptive Batch NormalizationSu Zheng, Zhengqi Gao, Fan-Keng Sun, Duane S. Boning 等NeurIPS 2024 · 被引用 5 次
- CFO: Learning Continuous-Time PDE Dynamics via Flow-Matched Neural OperatorsXianglong Hou, Xinquan Huang, Paris PerdikarisICLR 2026 · 被引用 10 次
