Stationary Activations for Uncertainty Calibration in Deep Learning
Lassi Meronen, Christabella Irwanto, Arno Solin
摘要
We introduce a new family of non-linear neural network activation functions that mimic the properties induced by the widely-used Matérn family of kernels in Gaussian process (GP) models. This class spans a range of locally stationary models of various degrees of mean-square differentiability. We show an explicit link to the corresponding GP models in the case that the network consists of one infinitely wide hidden layer. In the limit of infinite smoothness the Matérn family results in the RBF kernel, and in this case we recover RBF activations. Matérn activation functions result in similar appealing properties to their counterparts in GP models, and we demonstrate that the local stationarity property together with limited mean-square differentiability shows both good performance and uncertainty calibration in Bayesian deep learning tasks. In particular, local stationarity helps calibrate out-of-distribution (OOD) uncertainty. We demonstrate these properties on classification and regression benchmarks and a radar emitter classification task. ArcCos-0 kernel [8] ArcCos-1 kernel[8] ERF-NN kernel [68] RBF-NN kernel [68] Matérn-5 2 kernel MLP with step activation ReLU activation ERF (sigmoidal) activation RBF activation Matérn-5 2 activation (this paper) ArcCos-1 kernel [8] RBF-NN [68] Matérn-5 2 kernel Matérn-3 2 kernel Matérn-1 2 kernel MLP with ReLU activation RBF activation Matérn-5 2 activation Matérn-3 2 activation Matérn-1 2 activation
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- A Stitch in Time Saves Nine: A Train-Time Regularizing Loss for Improved Neural Network CalibrationRamya Hebbalaguppe, Jatin Prakash, Neelabh Madan, Chetan AroraCVPR 2022 · 被引用 38 次
- Deep Neural Networks as Point Estimates for Deep Gaussian ProcessesVincent Dutordoir, James Hensman, Mark van der Wilk, Carl Henrik Ek 等NeurIPS 2021 · 被引用 35 次
- Periodic Activation Functions Induce StationarityLassi Meronen, Martin Trapp, Arno SolinNeurIPS 2021 · 被引用 31 次
- Beyond Unimodal: Generalising Neural Processes for Multimodal Uncertainty EstimationMyong Chol Jung, He Zhao, Joanna Dipnall, Lan DuNeurIPS 2023 · 被引用 18 次
- Squared Neural Families: A New Class of Tractable Density ModelsRussell Tsuchida, Cheng Soon Ong, Dino SejdinovicNeurIPS 2023 · 被引用 15 次
它引用的顶会 Paper3
- Implicit Neural Representations with Periodic Activation FunctionsVincent Sitzmann, Julien N. P. Martel, Alexander W. Bergman, David B. Lindell 等NeurIPS 2020 · 被引用 4,008 次
- How Good is the Bayes Posterior in Deep Neural Networks Really?Florian Wenzel, Kevin Roth, Bastiaan S. Veeling, Jakub Swiatkowski 等ICML 2020 · 被引用 409 次
- On the Expressiveness of Approximate Inference in Bayesian Neural NetworksAndrew Y. K. Foong, David R. Burt, Yingzhen Li, Richard E. TurnerNeurIPS 2020 · 被引用 142 次
相关 Paper
- Activation-level uncertainty in deep neural networksPablo Morales-Alvarez, Daniel Hernández-Lobato, Rafael Molina, José Miguel Hernández-LobatoICLR 2021 · 被引用 16 次
- Avoiding Kernel Fixed Points: Computing with ELU and GELU Infinite NetworksRussell Tsuchida, Tim Pearce, Christopher van der Heide, Fred Roosta 等AAAI 2021 · 被引用 10 次
- Deep Kernel Posterior Learning under Infinite Variance Prior WeightsJorge Loría, Anindya BhadraICLR 2025
- Exploring the Uncertainty Properties of Neural Networks' Implicit Priors in the Infinite-Width LimitBen Adlam, Jaehoon Lee, Lechao Xiao, Jeffrey Pennington 等ICLR 2021 · 被引用 3 次
- Bayesian Deep Ensembles via the Neural Tangent KernelBobby He, Balaji Lakshminarayanan, Yee Whye TehNeurIPS 2020 · 被引用 136 次
