MCMC Should Mix: Learning Energy-Based Model with Neural Transport Latent Space MCMC
Erik Nijkamp, Ruiqi Gao, Pavel Sountsov, Srinivas Vasudevan, Bo Pang, Song-Chun Zhu, Ying Nian Wu
摘要
Learning energy-based model (EBM) requires MCMC sampling of the learned model as an inner loop of the learning algorithm. However, MCMC sampling of EBMs in high-dimensional data space is generally not mixing, because the energy function, which is usually parametrized by a deep network, is highly multi-modal in the data space. This is a serious handicap for both theory and practice of EBMs. In this paper, we propose to learn an EBM with a flow-based model (or in general a latent variable model) serving as a backbone, so that the EBM is a correction or an exponential tilting of the flow-based model. We show that the model has a particularly simple form in the space of the latent variables of the backbone model, and MCMC sampling of the EBM in the latent space mixes well and traverses modes in the data space. This enables proper sampling and learning of EBMs. * Equal contribution. † Majority of research was conducted at Google.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- EM Distillation for One-step Diffusion ModelsSirui Xie, Zhisheng Xiao, Diederik P. Kingma, Tingbo Hou 等NeurIPS 2024 · 被引用 69 次
- End-to-end Stochastic Optimization with Energy-based ModelLingkai Kong, Jiaming Cui, Yuchen Zhuang, Rui Feng 等NeurIPS 2022 · 被引用 33 次
- Energy-guided Entropic Neural Optimal TransportPetr Mokrov, Alexander Korotin, Alexander Kolesov, Nikita Gushchin 等ICLR 2024 · 被引用 30 次
- Energy-Based Models for Anomaly Detection: A Manifold Diffusion Recovery ApproachSangwoong Yoon, Young-Uk Jin, Yung-Kyun Noh, Frank C. ParkNeurIPS 2023 · 被引用 28 次
- On Sampling with Approximate Transport MapsLouis Grenioux, Alain Oliviero Durmus, Eric Moulines, Marylou GabriéICML 2023 · 被引用 25 次
它引用的顶会 Paper4
- Improved Contrastive Divergence Training of Energy-Based ModelsYilun Du, Shuang Li, Joshua B. Tenenbaum, Igor MordatchICML 2021 · 被引用 171 次
- VAEBM: A Symbiosis between Variational Autoencoders and Energy-based ModelsZhisheng Xiao, Karsten Kreis, Jan Kautz, Arash VahdatICLR 2021 · 被引用 139 次
- No MCMC for me: Amortized sampling for fast and stable training of energy-based modelsWill Sussman Grathwohl, Jacob Jin Kelly, Milad Hashemi, Mohammad Norouzi 等ICLR 2021 · 被引用 75 次
- Flow Contrastive Estimation of Energy-Based ModelsRuiqi Gao, Erik Nijkamp, Diederik P. Kingma, Zhen Xu 等CVPR 2020
相关 Paper
- Learning Energy-Based Prior Model with Diffusion-Amortized MCMCPeiyu Yu, Yaxuan Zhu, Sirui Xie, Xiaojian Ma 等NeurIPS 2023 · 被引用 17 次
- Generative Flow Networks for Discrete Probabilistic ModelingDinghuai Zhang, Nikolay Malkin, Zhen Liu, Alexandra Volokhova 等ICML 2022 · 被引用 131 次
- Learning Joint Latent Space EBM Prior Model for Multi-layer GeneratorJiali Cui, Ying Nian Wu, Tian HanCVPR 2023
- Learning Hierarchical Features with Joint Latent Space Energy-Based PriorJiali Cui, Ying Nian Wu, Tian HanICCV 2023 · 被引用 11 次
- Generalized Energy Based ModelsMichael Arbel, Liang Zhou, Arthur GrettonICLR 2021 · 被引用 254 次
