Learning Deep Latent Variable Models by Short-Run MCMC Inference With Optimal Transport Correction
Dongsheng An, Jianwen Xie, Ping Li
Abstract
Learning latent variable models with deep top-down architectures typically requires inferring the latent variables for each training example based on the posterior distribution of these latent variables. The inference step typically relies on either time-consuming long-run Markov chain Monte Carlo (MCMC) sampling or a separate inference model for variational learning. In this paper, we propose to use a shortrun MCMC, such as a short-run Langevin dynamics, as an approximate flow-based inference engine. The bias existing in the output distribution of the non-convergent short-run Langevin dynamics is corrected by the optimal transport (OT), which aims at transforming the biased distribution produced by the finite-step MCMC to the prior distribution with a minimum transport cost. Our experiments not only verify the effectiveness of the OT correction for the short-run MCMC, but also demonstrate that the latent variable model trained by the proposed strategy performs better than the variational auto-encoder (VAE) in terms of image reconstruction/generation and anomaly detection.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 83447c67-aefe-414e-be1b-60d15f2e98ddCited by top-tier papers4
- A Tale of Two Flows: Cooperative Learning of Langevin Flow and Normalizing Flow Toward Energy-Based ModelJianwen Xie, Yaxuan Zhu, Jun Li, Ping LiICLR 2022 · 53 citations
- Learning Unnormalized Statistical Models via Compositional OptimizationWei Jiang, Jiayu Qin, Lingyu Wu, Changyou Chen et al.ICML 2023 · 8 citations
- A Tale of Two Latent Flows: Learning Latent Space Normalizing Flow with Short-Run Langevin Flow for Approximate InferenceJianwen Xie, Yaxuan Zhu, Yifei Xu, Dingcheng Li et al.AAAI 2023 · 6 citations
- Learning Joint Latent Space EBM Prior Model for Multi-layer GeneratorJiali Cui, Ying Nian Wu, Tian HanCVPR 2023
Builds on9
- From Variational to Deterministic AutoencodersPartha Ghosh, Mehdi S. M. Sajjadi, Antonio Vergari, Michael J. Black et al.ICLR 2020 · 298 citations
- Learning Latent Space Energy-Based Prior ModelBo Pang, Tian Han, Erik Nijkamp, Song-Chun Zhu et al.NeurIPS 2020 · 152 citations
- Learning Feature-to-Feature Translator by Alternating Back-Propagation for Generative Zero-Shot LearningYizhe Zhu, Jianwen Xie, Bingchen Liu, Ahmed ElgammalICCV 2019 · 98 citations
- Recover and Identify: A Generative Dual Model for Cross-Resolution Person Re-IdentificationYu-Jhe Li, Yun-Chun Chen, Yen-Yu Lin, Xiaofei Du et al.ICCV 2019 · 88 citations
- Ae-OT: a New Generative Model based on Extended Semi-discrete Optimal transportDongsheng An, Yang Guo, Na Lei, Zhongxuan Luo et al.ICLR 2020 · 68 citations
Related papers
- Langevin Autoencoders for Learning Deep Latent Variable ModelsShohei Taniguchi, Yusuke Iwasawa, Wataru Kumagai, Yutaka MatsuoNeurIPS 2022 · 2 citations
- Black-Box Variational Inference as a Parametric Approximation to Langevin DynamicsMatthew D. Hoffman, Yian MaICML 2020 · 16 citations
- Generalized Variational Inference via Optimal TransportJinjin Chi, Zhichao Zhang, Zhiyao Yang, Jihong Ouyang et al.AAAI 2024 · 1 citation
- Learning Energy-Based Model with Variational Auto-Encoder as Amortized SamplerJianwen Xie, Zilong Zheng, Ping LiAAAI 2021 · 57 citations
- Coupled Variational AutoencoderXiaoran Hao, Patrick ShaftoICML 2023 · 7 citations
