Quasi-Monte Carlo Methods Enable Extremely Low-Dimensional Deep Generative Models
Miles Martinez, Alex H. Williams
Abstract
This paper introduces quasi-Monte Carlo latent variable models (QLVMs): a class of deep generative models that are specialized for finding extremely low-dimensional and interpretable embeddings of high-dimensional datasets. Unlike standard approaches, which rely on a learned encoder and variational lower bounds, QLVMs directly approximate the marginal likelihood by randomized quasi-Monte Carlo integration. While this brute force approach has drawbacks in higher-dimensional spaces, we find that it excels in fitting one, two, and three dimensional deep latent variable models. Empirical results on a range of datasets show that QLVMs consistently outperform conventional variational autoencoders (VAEs) and importance weighted autoencoders (IWAEs) with matched latent dimensionality. The resulting embeddings enable transparent visualization and post hoc analyses such as nonparametric density estimation, clustering, and geodesic path computation, which are nontrivial to validate in higher-dimensional spaces. While our approach is compute-intensive and struggles to generate fine-scale details in complex datasets, it offers a compelling solution for applications prioritizing interpretability and latent space analysis.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers1
Ask how each one uses itBuilds on8
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- NVAE: A Deep Hierarchical Variational AutoencoderArash Vahdat, Jan KautzNeurIPS 2020 · 1,141 citations
- The Usual Suspects? Reassessing Blame for VAE Posterior CollapseBin Dai, Ziyu Wang, David P. WipfICML 2020 · 89 citations
- Monte Carlo Variational Auto-EncodersAchille Thin, Nikita Kotelevskii, Arnaud Doucet, Alain Durmus et al.ICML 2021 · 51 citations
- Efficient Iterative Amortized Inference for Learning Symmetric and Disentangled Multi-Object RepresentationsPatrick Emami, Pan He, Sanjay Ranka, Anand RangarajanICML 2021 · 48 citations
Related papers
- Topic-VQ-VAE: Leveraging Latent Codebooks for Flexible Topic-Guided Document GenerationYoungjoon Yoo, Jongwon ChoiAAAI 2024 · 8 citations
- Crafting Interpretable Embeddings for Language Neuroscience by Asking LLMs QuestionsVinamra Benara, Chandan Singh, John X. Morris, Richard J. Antonello et al.NeurIPS 2024 · 26 citations
- Variational Pólya TreeLu Xu, Tsai Hor Chan, Lequan Yu, Kwok Fai Lam et al.NeurIPS 2025
- Langevin Autoencoders for Learning Deep Latent Variable ModelsShohei Taniguchi, Yusuke Iwasawa, Wataru Kumagai, Yutaka MatsuoNeurIPS 2022 · 2 citations
- Bayesian Regularization of Latent RepresentationChukwudi Paul Obite, Zhi Chang, Keyan Wu, Shiwei LanICLR 2025
