Learning Deep Latent Variable Models by Short-Run MCMC Inference With Optimal Transport Correction
Dongsheng An, Jianwen Xie, Ping Li
摘要
Learning latent variable models with deep top-down architectures typically requires inferring the latent variables for each training example based on the posterior distribution of these latent variables. The inference step typically relies on either time-consuming long-run Markov chain Monte Carlo (MCMC) sampling or a separate inference model for variational learning. In this paper, we propose to use a shortrun MCMC, such as a short-run Langevin dynamics, as an approximate flow-based inference engine. The bias existing in the output distribution of the non-convergent short-run Langevin dynamics is corrected by the optimal transport (OT), which aims at transforming the biased distribution produced by the finite-step MCMC to the prior distribution with a minimum transport cost. Our experiments not only verify the effectiveness of the OT correction for the short-run MCMC, but also demonstrate that the latent variable model trained by the proposed strategy performs better than the variational auto-encoder (VAE) in terms of image reconstruction/generation and anomaly detection.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- A Tale of Two Flows: Cooperative Learning of Langevin Flow and Normalizing Flow Toward Energy-Based ModelJianwen Xie, Yaxuan Zhu, Jun Li, Ping LiICLR 2022 · 被引用 53 次
- Learning Unnormalized Statistical Models via Compositional OptimizationWei Jiang, Jiayu Qin, Lingyu Wu, Changyou Chen 等ICML 2023 · 被引用 8 次
- A Tale of Two Latent Flows: Learning Latent Space Normalizing Flow with Short-Run Langevin Flow for Approximate InferenceJianwen Xie, Yaxuan Zhu, Yifei Xu, Dingcheng Li 等AAAI 2023 · 被引用 6 次
- Learning Joint Latent Space EBM Prior Model for Multi-layer GeneratorJiali Cui, Ying Nian Wu, Tian HanCVPR 2023
它引用的顶会 Paper9
- From Variational to Deterministic AutoencodersPartha Ghosh, Mehdi S. M. Sajjadi, Antonio Vergari, Michael J. Black 等ICLR 2020 · 被引用 298 次
- Learning Latent Space Energy-Based Prior ModelBo Pang, Tian Han, Erik Nijkamp, Song-Chun Zhu 等NeurIPS 2020 · 被引用 152 次
- Learning Feature-to-Feature Translator by Alternating Back-Propagation for Generative Zero-Shot LearningYizhe Zhu, Jianwen Xie, Bingchen Liu, Ahmed ElgammalICCV 2019 · 被引用 98 次
- Recover and Identify: A Generative Dual Model for Cross-Resolution Person Re-IdentificationYu-Jhe Li, Yun-Chun Chen, Yen-Yu Lin, Xiaofei Du 等ICCV 2019 · 被引用 88 次
- Ae-OT: a New Generative Model based on Extended Semi-discrete Optimal transportDongsheng An, Yang Guo, Na Lei, Zhongxuan Luo 等ICLR 2020 · 被引用 68 次
相关 Paper
- Langevin Autoencoders for Learning Deep Latent Variable ModelsShohei Taniguchi, Yusuke Iwasawa, Wataru Kumagai, Yutaka MatsuoNeurIPS 2022 · 被引用 2 次
- Black-Box Variational Inference as a Parametric Approximation to Langevin DynamicsMatthew D. Hoffman, Yian MaICML 2020 · 被引用 16 次
- Generalized Variational Inference via Optimal TransportJinjin Chi, Zhichao Zhang, Zhiyao Yang, Jihong Ouyang 等AAAI 2024 · 被引用 1 次
- Learning Energy-Based Model with Variational Auto-Encoder as Amortized SamplerJianwen Xie, Zilong Zheng, Ping LiAAAI 2021 · 被引用 57 次
- Coupled Variational AutoencoderXiaoran Hao, Patrick ShaftoICML 2023 · 被引用 7 次
