Entropy-based Training Methods for Scalable Neural Implicit Samplers
Weijian Luo, Boya Zhang, Zhihua Zhang
摘要
Efficiently sampling from un-normalized target distributions is a fundamental problem in scientific computing and machine learning. Traditional approaches such as Markov Chain Monte Carlo (MCMC) guarantee asymptotically unbiased samples from such distributions but suffer from computational inefficiency, particularly when dealing with high-dimensional targets, as they require numerous iterations to generate a batch of samples. In this paper, we introduce an efficient and scalable neural implicit sampler that overcomes these limitations. The implicit sampler can generate large batches of samples with low computational costs by leveraging a neural transformation that directly maps easily sampled latent vectors to target samples without the need for iterative procedures. To train the neural implicit samplers, we introduce two novel methods: the KL training method and the Fisher training method. The former method minimizes the Kullback-Leibler divergence, while the latter minimizes the Fisher divergence between the sampler and the target distributions. By employing the two training methods, we effectively optimize the neural implicit samplers to learn and generate from the desired target distribution. To demonstrate the effectiveness, efficiency, and scalability of our proposed samplers, we evaluate them on three sampling benchmarks with different scales. These benchmarks include sampling from 2D targets, Bayesian inference, and sampling from high-dimensional energy-based models (EBMs). Notably, in the experiment involving high-dimensional EBMs, our sampler produces samples that are comparable to those generated by MCMC-based methods while being more than 100 times more efficient, showcasing the efficiency of our neural sampler. Besides the theoretical contributions and strong empirical performances, the proposed neural samplers and corresponding training methods will shed light on further research on developing efficient samplers for various applications beyond the ones explored in this study.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- One-Step Diffusion Distillation through Score Implicit MatchingWeijian Luo, Zemin Huang, Zhengyang Geng, J. Zico Kolter 等NeurIPS 2024 · 被引用 81 次
- Vision-Language Navigation with Energy-Based PolicyRui Liu, Wenguan Wang, Yi YangNeurIPS 2024 · 被引用 39 次
- Schedule On the Fly: Diffusion Time Prediction for Faster and Better Image GenerationZilyu Ye, Zhiyang Chen, Tiancheng Li, Zemin Huang 等CVPR 2025
- David and Goliath: Small One-step Model Beats Large Diffusion with Score Post-trainingWeijian Luo, Colin Zhang, Debing Zhang, Zhengyang GengICML 2025
它引用的顶会 Paper23
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li 等NeurIPS 2022 · 被引用 8,965 次
- Improved Denoising Diffusion Probabilistic ModelsAlexander Quinn Nichol, Prafulla DhariwalICML 2021 · 被引用 5,234 次
- GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion ModelsAlexander Quinn Nichol, Prafulla Dhariwal, Aditya Ramesh, Pranav Shyam 等ICML 2022 · 被引用 4,691 次
相关 Paper
- Revisiting Unbiased Implicit Variational InferenceTobias Pielok, Bernd Bischl, David RügamerICML 2025
- Hamiltonian Dynamics with Non-Newtonian Momentum for Rapid SamplingGreg Ver Steeg, Aram GalstyanNeurIPS 2021 · 被引用 18 次
- Generative Particle Variational Inference via Estimation of Functional GradientsNeale Ratzlaff, Qinxun Bai, Fuxin Li, Wei XuICML 2021
- Implicit Variational Inference for High-Dimensional PosteriorsAnshuk Uppal, Kristoffer Stensbo-Smidt, Wouter Boomsma, Jes FrellsenNeurIPS 2023 · 被引用 6 次
- Annealing Flow Generative Models Towards Sampling High-Dimensional and Multi-Modal DistributionsDongze Wu, Yao XieICML 2025
