Being Bayesian about Categorical Probability
Taejong Joo, Uijung Chung, Min-Gwan Seo
Abstract
Neural networks utilize the softmax as a building block in classification tasks, which contains an overconfidence problem and lacks an uncertainty representation ability. As a Bayesian alternative to the softmax, we consider a random variable of a categorical probability over class labels. In this framework, the prior distribution explicitly models the presumed noise inherent in the observed label, which provides consistent gains in generalization performance in multiple challenging tasks. The proposed method inherits advantages of Bayesian approaches that achieve better uncertainty estimation and model calibration. Our method can be implemented as a plug-and-play loss function with negligible computational overhead compared to the softmax with the cross-entropy loss function.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f2ff6ce8-72ab-4627-a5b2-73d32aabdf97Cited by top-tier papers17
- Deep Evidential RegressionAlexander Amini, Wilko Schwarting, Ava Soleimany, Daniela RusNeurIPS 2020 · 777 citations
- Better Uncertainty Calibration via Proper Scores for Classification and BeyondSebastian G. Gruber, Florian BuettnerNeurIPS 2022 · 88 citations
- Post-hoc Uncertainty Learning Using a Dirichlet Meta-ModelMaohao Shen, Yuheng Bu, Prasanna Sattigeri, Soumya Ghosh et al.AAAI 2023 · 51 citations
- Towards Trustworthy Predictions from Deep Neural Networks with Fast Adversarial CalibrationChristian Tomani, Florian BuettnerAAAI 2021 · 43 citations
- Are Uncertainty Quantification Capabilities of Evidential Deep Learning a Mirage?Maohao Shen, Jongha Jon Ryu, Soumya Ghosh, Yuheng Bu et al.NeurIPS 2024 · 29 citations
Builds on1
Related papers
- Energy-Based Open-World Uncertainty Modeling for Confidence CalibrationYezhen Wang, Bo Li, Tong Che, Kaiyang Zhou et al.ICCV 2021 · 78 citations
- Exploring the Uncertainty Properties of Neural Networks' Implicit Priors in the Infinite-Width LimitBen Adlam, Jaehoon Lee, Lechao Xiao, Jeffrey Pennington et al.ICLR 2021 · 3 citations
- Rethinking Approximate Gaussian Inference in ClassificationBálint Mucsányi, Nathaël Da Costa, Philipp HennigNeurIPS 2025 · 2 citations
- On Uncertainty, Tempering, and Data Augmentation in Bayesian ClassificationSanyam Kapoor, Wesley J. Maddox, Pavel Izmailov, Andrew Gordon WilsonNeurIPS 2022 · 64 citations
- Conservative Uncertainty Estimation By Fitting Prior NetworksKamil Ciosek, Vincent Fortuin, Ryota Tomioka, Katja Hofmann et al.ICLR 2020 · 65 citations
