Learning Energy-Based Generative Models via Coarse-to-Fine Expanding and Sampling
Yang Zhao, Jianwen Xie, Ping Li
摘要
Energy-based models (EBMs) parameterized by neural networks can be trained by the Markov chain Monte Carlo (MCMC) sampling-based maximum likelihood estimation. Despite the recent significant success of EBMs in image generation, the current approaches to train EBMs are unstable and have difficulty synthesizing diverse and high-fidelity images. In this paper, we propose to train EBMs via a multistage coarse-to-fine expanding and sampling strategy, which starts with learning a coarse-level EBM from images at low resolution and then gradually transits to learn a finer-level EBM from images at higher resolution by expanding the energy function as the learning progresses. The proposed framework is computationally efficient with smooth learning and sampling. It achieves the best performance on image generation amongst all EBMs and is the first successful EBM to synthesize high-fidelity images at 512 × 512 resolution. It can also be useful for image restoration and out-of-distribution detection. Lastly, the proposed framework is further generalized to the one-sided unsupervised image-to-image translation and beats baseline methods in terms of model size and training budget. We also present a gradient-based generative saliency method to interpret the translation dynamics.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper25
- COLD Decoding: Energy-based Constrained Text Generation with Langevin DynamicsLianhui Qin, Sean Welleck, Daniel Khashabi, Yejin ChoiNeurIPS 2022 · 被引用 217 次
- Learning Generative Vision Transformer with Energy-Based Latent Space for Saliency PredictionJing Zhang, Jianwen Xie, Nick Barnes, Ping LiNeurIPS 2021 · 被引用 117 次
- A Tale of Two Flows: Cooperative Learning of Langevin Flow and Normalizing Flow Toward Energy-Based ModelJianwen Xie, Yaxuan Zhu, Jun Li, Ping LiICLR 2022 · 被引用 53 次
- Energy-based Latent Aligner for Incremental LearningK. J. Joseph, Salman Khan, Fahad Shahbaz Khan, Rao Muhammad Anwer 等CVPR 2022 · 被引用 35 次
- Energy-guided Entropic Neural Optimal TransportPetr Mokrov, Alexander Korotin, Alexander Kolesov, Nikita Gushchin 等ICLR 2024 · 被引用 30 次
它引用的顶会 Paper14
- Training Generative Adversarial Networks with Limited DataTero Karras, Miika Aittala, Janne Hellsten, Samuli Laine 等NeurIPS 2020 · 被引用 2,345 次
- Energy-based Out-of-distribution DetectionWeitang Liu, Xiaoyun Wang, John D. Owens, Yixuan LiNeurIPS 2020 · 被引用 2,213 次
- Improved Techniques for Training Score-Based Generative ModelsYang Song, Stefano ErmonNeurIPS 2020 · 被引用 1,527 次
- Your classifier is secretly an energy based model and you should treat it like oneWill Grathwohl, Kuan-Chieh Wang, Jörn-Henrik Jacobsen, David Duvenaud 等ICLR 2020 · 被引用 643 次
- U-GAT-IT: Unsupervised Generative Attentional Networks with Adaptive Layer-Instance Normalization for Image-to-Image TranslationJunho Kim, Minjae Kim, Hyeonwoo Kang, Kwanghee LeeICLR 2020 · 被引用 632 次
相关 Paper
- Learning Energy-Based Models by Diffusion Recovery LikelihoodRuiqi Gao, Yang Song, Ben Poole, Ying Nian Wu 等ICLR 2021 · 被引用 144 次
- Learning Energy-Based Models by Cooperative Diffusion Recovery LikelihoodYaxuan Zhu, Jianwen Xie, Ying Nian Wu, Ruiqi GaoICLR 2024 · 被引用 18 次
- Bi-level Doubly Variational Learning for Energy-based Latent Variable ModelsGe Kan, Jinhu Lü, Tian Wang, Baochang Zhang 等CVPR 2022 · 被引用 4 次
- VAEBM: A Symbiosis between Variational Autoencoders and Energy-based ModelsZhisheng Xiao, Karsten Kreis, Jan Kautz, Arash VahdatICLR 2021 · 被引用 139 次
- Energy-Inspired Self-Supervised Pretraining for Vision ModelsZe Wang, Jiang Wang, Zicheng Liu, Qiang QiuICLR 2023 · 被引用 1 次
