DIBS: Diversity Inducing Information Bottleneck in Model Ensembles
Samarth Sinha, Homanga Bharadhwaj, Anirudh Goyal, Hugo Larochelle, Animesh Garg, Florian Shkurti
Abstract
Although deep learning models have achieved state-of-the art performance on a number of vision tasks, generalization over high dimensional multi-modal data, and reliable predictive uncertainty estimation are still active areas of research. Bayesian approaches including Bayesian Neural Nets (BNNs) do not scale well to modern computer vision tasks, as they are difficult to train, and have poor generalization under dataset-shift (Lakshminarayanan, Pritzel, and Blundell 2017; Ovadia et al. 2019) . This motivates the need for effective ensembles which can generalize and give reliable uncertainty estimates. In this paper, we target the problem of generating effective ensembles of neural networks by encouraging diversity in prediction. We explicitly optimize a diversity inducing adversarial loss for learning the stochastic latent variables and thereby obtain diversity in the output predictions necessary for modeling multi-modal data. We evaluate our method on benchmark datasets: MNIST, CIFAR100, TinyIm-ageNet and MIT Places 2, and compared to the most competitive baselines show significant improvements in classification accuracy, under a shift in the data distribution and in out-of-distribution detection. : over 10% relative improvement in classification accuracy, over 5% relative improvement in generalizing under dataset shift, and over 5% better predictive uncertainty estimation as inferred by efficient outof-distribution (OOD) detection.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6b16e412-28b0-4efd-ba31-14fadeb5137eCited by top-tier papers7
- DVERGE: Diversifying Vulnerabilities for Enhanced Robust Generation of EnsemblesHuanrui Yang, Jingyang Zhang, Hongliang Dong, Nathan Inkawhich et al.NeurIPS 2020 · 144 citations
- DICE: Diversity in Deep Ensembles via Conditional Redundancy Adversarial EstimationAlexandre Ramé, Matthieu CordICLR 2021 · 60 citations
- On Power Laws in Deep EnsemblesEkaterina Lobacheva, Nadezhda Chirkova, Maxim Kodryan, Dmitry P. VetrovNeurIPS 2020 · 48 citations
- Fantastic Gains and Where to Find Them: On the Existence and Prospect of General Knowledge Transfer between Any Pretrained ModelKarsten Roth, Lukas Thede, A. Sophia Koepke, Oriol Vinyals et al.ICLR 2024 · 17 citations
- Agree to Disagree: Diversity through Disagreement for Better TransferabilityMatteo Pagliardini, Martin Jaggi, François Fleuret, Sai Praneeth KarimireddyICLR 2023 · 7 citations
Builds on1
Related papers
- Maximizing Overall Diversity for Improved Uncertainty Estimates in Deep EnsemblesSiddhartha Jain, Ge Liu, Jonas Mueller, David GiffordAAAI 2020 · 69 citations
- Credal Wrapper of Model Averaging for Uncertainty Estimation in ClassificationKaizheng Wang, Fabio Cuzzolin, Keivan Shariatmadar, David Moens et al.ICLR 2025
- Microcanonical Langevin Ensembles: Advancing the Sampling of Bayesian Neural NetworksEmanuel Sommer, Jakob Robnik, Giorgi Nozadze, Uros Seljak et al.ICLR 2025
- Neural Ensemble Search for Uncertainty Estimation and Dataset ShiftSheheryar Zaidi, Arber Zela, Thomas Elsken, Chris C. Holmes et al.NeurIPS 2021 · 97 citations
- Uncertainty-Aware Deep Classifiers Using Generative ModelsMurat Sensoy, Lance M. Kaplan, Federico Cerutti, Maryam SalekiAAAI 2020 · 88 citations
