Hamiltonian Monte Carlo on ReLU Neural Networks is Inefficient
Vu C. Dinh, Lam Si Tung Ho, Cuong V. Nguyen
摘要
We analyze the error rates of the Hamiltonian Monte Carlo algorithm with leapfrog integrator for Bayesian neural network inference. We show that due to the non-differentiability of activation functions in the ReLU family, leapfrog HMC for networks with these activation functions has a large local error rate of rather than the classical error rate of . This leads to a higher rejection rate of the proposals, making the method inefficient. We then verify our theoretical findings through empirical simulations as well as experiments on a real-world dataset that highlight the inefficiency of HMC inference on ReLU-based neural networks compared to analytical networks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper5
- A mathematical model for automatic differentiation in machine learningJérôme Bolte, Edouard PauwelsNeurIPS 2020 · 被引用 84 次
- Numerical influence of ReLU'(0) on backpropagationDavid Bertoin, Jérôme Bolte, Sébastien Gerchinovitz, Edouard PauwelsNeurIPS 2021 · 被引用 30 次
- Mixed Hamiltonian Monte Carlo for Mixed Discrete and Continuous VariablesGuangyao ZhouNeurIPS 2020 · 被引用 25 次
- MCMC Should Mix: Learning Energy-Based Model with Neural Transport Latent Space MCMCErik Nijkamp, Ruiqi Gao, Pavel Sountsov, Srinivas Vasudevan 等ICLR 2022 · 被引用 25 次
- A Gradient Based Strategy for Hamiltonian Monte Carlo Hyperparameter OptimizationAndrew Campbell, Wenlong Chen, Vincent Stimper, José Miguel Hernández-Lobato 等ICML 2021 · 被引用 20 次
相关 Paper
- On the Expressiveness of Approximate Inference in Bayesian Neural NetworksAndrew Y. K. Foong, David R. Burt, Yingzhen Li, Richard E. TurnerNeurIPS 2020 · 被引用 142 次
- Structured Stochastic Gradient MCMCAntonios Alexos, Alex J. Boyd, Stephan MandtICML 2022 · 被引用 14 次
- What Are Bayesian Neural Network Posteriors Really Like?Pavel Izmailov, Sharad Vikram, Matthew D. Hoffman, Andrew Gordon WilsonICML 2021 · 被引用 458 次
- Liberty or Depth: Deep Bayesian Neural Nets Do Not Need Complex Weight Posterior ApproximationsSebastian Farquhar, Lewis Smith, Yarin GalNeurIPS 2020 · 被引用 47 次
- Posterior Refinement Improves Sample Efficiency in Bayesian Neural NetworksAgustinus Kristiadi, Runa Eschenhagen, Philipp HennigNeurIPS 2022 · 被引用 17 次
