Model Selection for Bayesian Autoencoders
Ba-Hien Tran, Simone Rossi, Dimitrios Milios, Pietro Michiardi, Edwin V. Bonilla, Maurizio Filippone
Abstract
We develop a novel method for carrying out model selection for Bayesian autoencoders (BAEs) by means of prior hyper-parameter optimization. Inspired by the common practice of type-II maximum likelihood optimization and its equivalence to Kullback-Leibler divergence minimization, we propose to optimize the distributional sliced-Wasserstein distance (DSWD) between the output of the autoencoder and the empirical data distribution. The advantages of this formulation are that we can estimate the DSWD based on samples and handle high-dimensional problems. We carry out posterior estimation of the BAE parameters via stochastic gradient Hamiltonian Monte Carlo and turn our BAE into a generative model by fitting a flexible Dirichlet mixture model in the latent space. Consequently, we obtain a powerful alternative to variational autoencoders, which are the preferred choice in modern applications of autoencoders for representation learning with uncertainty. We evaluate our approach qualitatively and quantitatively using a vast experimental campaign on a number of unsupervised learning tasks and show that, in small-data regimes where priors matter, our approach provides state-of-the-art results, outperforming multiple competitive baselines.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9326cee8-112b-481e-9c0d-0688dfc404dcCited by top-tier papers6
- Fully Bayesian Autoencoders with Latent Sparse Gaussian ProcessesBa-Hien Tran, Babak Shahbaba, Stephan Mandt, Maurizio FilipponeICML 2023 · 9 citations
- On permutation symmetries in Bayesian neural network posteriors: a variational perspectiveSimone Rossi, Ankit Singh, Thomas HannaganNeurIPS 2023 · 5 citations
- One-Line-of-Code Data Mollification Improves Optimization of Likelihood-based Generative ModelsBa-Hien Tran, Giulio Franzese, Pietro Michiardi, Maurizio FilipponeNeurIPS 2023 · 4 citations
- Variational Learning of Fractional PosteriorsKian Ming A. Chai, Edwin V. BonillaICML 2025
- Neighbour-Driven Gaussian Process Variational Autoencoders for Scalable Structured Latent ModellingXinxing Shi, Xiaoyu Jiang, Mauricio A. ÁlvarezICML 2025
Builds on10
- Bayesian Deep Learning and a Probabilistic Perspective of GeneralizationAndrew Gordon Wilson, Pavel IzmailovNeurIPS 2020 · 845 citations
- Differentiable Augmentation for Data-Efficient GAN TrainingShengyu Zhao, Zhijian Liu, Ji Lin, Jun-Yan Zhu et al.NeurIPS 2020 · 707 citations
- What Are Bayesian Neural Network Posteriors Really Like?Pavel Izmailov, Sharad Vikram, Matthew D. Hoffman, Andrew Gordon WilsonICML 2021 · 458 citations
- How Good is the Bayes Posterior in Deep Neural Networks Really?Florian Wenzel, Kevin Roth, Bastiaan S. Veeling, Jakub Swiatkowski et al.ICML 2020 · 409 citations
- From Variational to Deterministic AutoencodersPartha Ghosh, Mehdi S. M. Sajjadi, Antonio Vergari, Michael J. Black et al.ICLR 2020 · 298 citations
Related papers
- Gromov-Wasserstein AutoencodersNao Nakagawa, Ren Togo, Takahiro Ogawa, Miki HaseyamaICLR 2023 · 2 citations
- Laplacian Autoencoders for Learning Stochastic RepresentationsMarco Miani, Frederik Warburg, Pablo Moreno-Muñoz, Nicki Skafte Detlefsen et al.NeurIPS 2022 · 17 citations
- Undirected Graphical Models as Approximate PosteriorsArash Vahdat, Evgeny Andriyash, William G. MacreadyICML 2020 · 15 citations
- Meta-Learning with Shared Amortized Variational InferenceEkaterina Iakovleva, Jakob Verbeek, Karteek AlahariICML 2020 · 25 citations
- Efficient Sliced Wasserstein Distance Computation via Adaptive Bayesian OptimizationManish Acharya, David HydeICLR 2026
