Exemplar VAE: Linking Generative Models, Nearest Neighbor Retrieval, and Data Augmentation
Sajad Norouzi, David J. Fleet, Mohammad Norouzi
Abstract
We introduce Exemplar VAEs, a family of generative models that bridge the gap between parametric and non-parametric, exemplar based generative models. Exemplar VAE is a variant of VAE with a non-parametric prior in the latent space based on a Parzen window estimator. To sample from it, one first draws a random exemplar from a training set, then stochastically transforms that exemplar into a latent code and a new observation. We propose retrieval augmented training (RAT) as a way to speed up Exemplar VAE training by using approximate nearest neighbor search in the latent space to define a lower bound on log marginal likelihood. To enhance generalization, model parameters are learned using exemplar leave-oneout and subsampling. Experiments demonstrate the effectiveness of Exemplar VAEs on density estimation and representation learning. Importantly, generative data augmentation using Exemplar VAEs on permutation invariant MNIST and Fashion MNIST reduces classification error from 1.17% to 0.69% and from 8.56% to 8.16%. Code is available at https://github.com/sajadn/Exemplar-VAE .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b765ec46-1dda-47e5-9d91-553e44349a73Cited by top-tier papers3
- Prototypical Variational Autoencoder for 3D Few-shot Object DetectionWeiliang Tang, Biqi Yang, Xianzhi Li, Yun-Hui Liu et al.NeurIPS 2023 · 8 citations
- ByPE-VAE: Bayesian Pseudocoresets Exemplar VAEQingzhong Ai, Lirong He, Shiyu Liu, Zenglin XuNeurIPS 2021 · 4 citations
- Multi-Rate VAE: Train Once, Get the Full Rate-Distortion CurveJuhan Bae, Michael R. Zhang, Michael Ruan, Eric Wang et al.ICLR 2023 · 3 citations
Builds on3
- Generalization through Memorization: Nearest Neighbor Language ModelsUrvashi Khandelwal, Omer Levy, Dan Jurafsky, Luke Zettlemoyer et al.ICLR 2020 · 1,038 citations
- From Variational to Deterministic AutoencodersPartha Ghosh, Mehdi S. M. Sajjadi, Antonio Vergari, Michael J. Black et al.ICLR 2020 · 298 citations
- A Forest from the Trees: Generation through NeighborhoodsYang Li, Tianxiang Gao, Junier OlivaAAAI 2020 · 5 citations
Related papers
- RAE: A Neural Network Dimensionality Reduction Method for Nearest Neighbors Preservation in Vector SearchHan Zhang, Dongfang ZhaoKDD 2026
- Gradient Origin NetworksSam Bond-Taylor, Chris G. WillcocksICLR 2021 · 2 citations
- BooVAE: Boosting Approach for Continual Learning of VAEEvgenii Egorov, Anna Kuzina, Evgeny BurnaevNeurIPS 2021 · 34 citations
- A Bayesian Nonparametrics View into Deep RepresentationsMichal Jamroz, Marcin Kurdziel, Mateusz OpalaNeurIPS 2020 · 1 citation
- Cracking Vector Search IndexesVasilis Mageirakos, Bowen Wu, Gustavo AlonsoVLDB 2025 · 6 citations
