Learning Energy-Based Model with Variational Auto-Encoder as Amortized Sampler
Jianwen Xie, Zilong Zheng, Ping Li
Abstract
Due to the intractable partition function, training energy-based models (EBMs) by maximum likelihood requires Markov chain Monte Carlo (MCMC) sampling to approximate the gradient of the Kullback-Leibler divergence between data and model distributions. However, it is non-trivial to sample from an EBM because of the difficulty of mixing between modes. In this paper, we propose to learn a variational auto-encoder (VAE) to initialize the finite-step MCMC, such as Langevin dynamics that is derived from the energy function, for efficient amortized sampling of the EBM. With these amortized MCMC samples, the EBM can be trained by maximum likelihood, which follows an "analysis by synthesis" scheme; while the VAE learns from these MCMC samples via variational Bayes. We call this joint training algorithm the variational MCMC teaching, in which the VAE chases the EBM toward data distribution. We interpret the learning algorithm as a dynamic alternating projection in the context of information geometry. Our proposed models can generate samples comparable to GANs and EBMs. Additionally, we demonstrate that our model can learn effective probabilistic distribution toward supervised conditional learning tasks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d0be61ee-2a15-4666-978e-b69851c240f3Cited by top-tier papers30
- Reduce, Reuse, Recycle: Compositional Generation with Energy-Based Diffusion Models and MCMCYilun Du, Conor Durkan, Robin Strudel, Joshua B. Tenenbaum et al.ICML 2023 · 219 citations
- Learning Generative Vision Transformer with Energy-Based Latent Space for Saliency PredictionJing Zhang, Jianwen Xie, Nick Barnes, Ping LiNeurIPS 2021 · 117 citations
- Object Representations as Fixed Points: Training Iterative Refinement Algorithms with Implicit DifferentiationMichael Chang, Tom Griffiths, Sergey LevineNeurIPS 2022 · 69 citations
- A Tale of Two Flows: Cooperative Learning of Langevin Flow and Normalizing Flow Toward Energy-Based ModelJianwen Xie, Yaxuan Zhu, Jun Li, Ping LiICLR 2022 · 53 citations
- Learning Energy-Based Generative Models via Coarse-to-Fine Expanding and SamplingYang Zhao, Jianwen Xie, Ping LiICLR 2021 · 51 citations
Builds on3
- Your classifier is secretly an energy based model and you should treat it like oneWill Grathwohl, Kuan-Chieh Wang, Jörn-Henrik Jacobsen, David Duvenaud et al.ICLR 2020 · 643 citations
- Energy-based models for atomic-resolution protein conformationsYilun Du, Joshua Meier, Jerry Ma, Rob Fergus et al.ICLR 2020 · 60 citations
- Learning Cycle-Consistent Cooperative Networks via Alternating MCMC Teaching for Unsupervised Cross-Domain TranslationJianwen Xie, Zilong Zheng, Xiaolin Fang, Song-Chun Zhu et al.AAAI 2021 · 14 citations
Related papers
- Learning Energy-based Model via Dual-MCMC TeachingJiali Cui, Tian HanNeurIPS 2023 · 14 citations
- Langevin Autoencoders for Learning Deep Latent Variable ModelsShohei Taniguchi, Yusuke Iwasawa, Wataru Kumagai, Yutaka MatsuoNeurIPS 2022 · 2 citations
- Joint Training of Variational Auto-Encoder and Latent Energy-Based ModelTian Han, Erik Nijkamp, Linqi Zhou, Bo Pang et al.CVPR 2020
- Efficient Training of Energy-Based Models Using Jarzynski EqualityDavide Carbone, Mengjian Hua, Simon Coste, Eric Vanden-EijndenNeurIPS 2023 · 21 citations
- Latent-Guided Cooperative Energy-Based ModelsCong Geng, Xue Han, Ye Yuan, Qiang Hu et al.ICML 2026
