Latent Time Neural Ordinary Differential Equations
Srinivas Anumasa, P. K. Srijith
摘要
Neural ordinary differential equations (NODE) have been proposed as a continuous depth generalization to popular deep learning models such as Residual networks (ResNets). They provide parameter efficiency and automate the model selection process in deep learning models to some extent. However, they lack the much-required uncertainty modelling and robustness capabilities which are crucial for their use in several real-world applications such as autonomous driving and healthcare. We propose a novel and unique approach to model uncertainty in NODE by considering a distribution over the end-time T of the ODE solver. The proposed approach, latent time NODE (LT-NODE), treats T as a latent variable and apply Bayesian learning to obtain a posterior distribution over T from the data. In particular, we use variational inference to learn an approximate posterior and the model parameters. Prediction is done by considering the NODE representations from different samples of the posterior and can be done efficiently using a single forward pass. As T implicitly defines the depth of a NODE, posterior distribution over T would also help in model selection in NODE. We also propose, adaptive latent time NODE (ALT-NODE), which allow each data point to have a distinct posterior distribution over end-times. ALT-NODE uses amortized variational inference to learn an approximate posterior using inference networks. We demonstrate the effectiveness of the proposed approaches in modelling uncertainty and robustness through experiments on synthetic and several real-world image classification data.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Improving Transferability for Cross-Domain Trajectory Prediction via Neural Stochastic Differential EquationDaehee Park, Jaewoo Jeong, Kuk-Jin YoonAAAI 2024 · 被引用 17 次
- Learning Coupled Continuous-Time Latent Dynamics from Irregular EventsJiankai Zuo, Yang Zhang, Yu Zhang, Jiarui Liang 等ICML 2026
它引用的顶会 Paper8
- Hyperparameter Ensembles for Robustness and Uncertainty QuantificationFlorian Wenzel, Jasper Snoek, Dustin Tran, Rodolphe JenattonNeurIPS 2020 · 被引用 263 次
- Dissecting Neural ODEsStefano Massaroli, Michael Poli, Jinkyoo Park, Atsushi Yamashita 等NeurIPS 2020 · 被引用 261 次
- On Robustness of Neural Ordinary Differential EquationsHanshu Yan, Jiawei Du, Vincent Y. F. Tan, Jiashi FengICLR 2020 · 被引用 161 次
- SDE-Net: Equipping Deep Neural Networks with Uncertainty EstimatesLingkai Kong, Jimeng Sun, Chao ZhangICML 2020 · 被引用 134 次
- Adaptive Checkpoint Adjoint Method for Gradient Estimation in Neural ODEJuntang Zhuang, Nicha C. Dvornek, Xiaoxiao Li, Sekhar Tatikonda 等ICML 2020 · 被引用 125 次
相关 Paper
- Efficient Certified Training and Robustness Verification of Neural ODEsMustafa Zeqiri, Mark Niklas Müller, Marc Fischer, Martin T. VechevICLR 2023
- Imbedding Deep Neural NetworksAndrew Corbett, Dmitry KanginICLR 2022 · 被引用 2 次
- Improving Neural ODE Training with Temporal Adaptive Batch NormalizationSu Zheng, Zhengqi Gao, Fan-Keng Sun, Duane S. Boning 等NeurIPS 2024 · 被引用 5 次
- Neural Delay Differential EquationsQunxi Zhu, Yao Guo, Wei LinICLR 2021 · 被引用 3 次
- Balanced Neural ODEs: nonlinear model order reduction and Koopman operator approximationsJulius Aka, Johannes Brunnemann, Jörg Eiden, Arne Speerforck 等ICLR 2025 · 被引用 1 次
