A Langevin-like Sampler for Discrete Distributions
Ruqi Zhang, Xingchao Liu, Qiang Liu
摘要
We propose discrete Langevin proposal (DLP), a simple and scalable gradient-based proposal for sampling complex high-dimensional discrete distributions. In contrast to Gibbs sampling-based methods, DLP is able to update all coordinates in parallel in a single step and the magnitude of changes is controlled by a stepsize. This allows a cheap and efficient exploration in the space of high-dimensional and strongly correlated variables. We prove the efficiency of DLP by showing that the asymptotic bias of its stationary distribution is zero for log-quadratic distributions, and is small for distributions that are close to being log-quadratic. With DLP, we develop several variants of sampling algorithms, including unadjusted, Metropolis-adjusted, stochastic and preconditioned versions. DLP outperforms many popular alternatives on a wide variety of tasks, including Ising models, restricted Boltzmann machines, deep energy-based models, binary neural networks and language generation.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper26
- Automatically Auditing Large Language Models via Discrete OptimizationErik Jones, Anca D. Dragan, Aditi Raghunathan, Jacob SteinhardtICML 2023 · 被引用 232 次
- Generative Flow Networks for Discrete Probabilistic ModelingDinghuai Zhang, Nikolay Malkin, Zhen Liu, Alexandra Volokhova 等ICML 2022 · 被引用 131 次
- BayesDAG: Gradient-Based Posterior Inference for Causal DiscoveryYashas Annadani, Nick Pawlowski, Joel Jennings, Stefan Bauer 等NeurIPS 2023 · 被引用 54 次
- Revisiting Sampling for Combinatorial OptimizationHaoran Sun, Katayoon Goshvadi, Azade Nova, Dale Schuurmans 等ICML 2023 · 被引用 28 次
- Alleviating Adversarial Attacks on Variational Autoencoders with MCMCAnna Kuzina, Max Welling, Jakub M. TomczakNeurIPS 2022 · 被引用 16 次
它引用的顶会 Paper9
- Your classifier is secretly an energy based model and you should treat it like oneWill Grathwohl, Kuan-Chieh Wang, Jörn-Henrik Jacobsen, David Duvenaud 等ICLR 2020 · 被引用 643 次
- Cyclical Stochastic Gradient MCMC for Bayesian Deep LearningRuqi Zhang, Chunyuan Li, Jianyi Zhang, Changyou Chen 等ICLR 2020 · 被引用 292 次
- Understanding Knowledge Distillation in Non-autoregressive Machine TranslationChunting Zhou, Jiatao Gu, Graham NeubigICLR 2020 · 被引用 235 次
- Generative Flow Networks for Discrete Probabilistic ModelingDinghuai Zhang, Nikolay Malkin, Zhen Liu, Alexandra Volokhova 等ICML 2022 · 被引用 131 次
- Oops I Took A Gradient: Scalable Sampling for Discrete DistributionsWill Grathwohl, Kevin Swersky, Milad Hashemi, David Duvenaud 等ICML 2021 · 被引用 113 次
相关 Paper
- Optimal Scaling for Locally Balanced Proposals in Discrete SpacesHaoran Sun, Hanjun Dai, Dale SchuurmansNeurIPS 2022 · 被引用 14 次
- Exploring Non-Convex Discrete Energy Landscapes: An Efficient Langevin-Like Sampler with Replica ExchangeHaoyang Zheng, Hengrong Du, Ruqi Zhang, Guang LinAAAI 2026
- Optimal Underdamped Langevin MCMC MethodZhengmian Hu, Feihu Huang, Heng HuangNeurIPS 2021 · 被引用 5 次
- LSB: Local Self-Balancing MCMC in Discrete SpacesEmanuele SansoneICML 2022 · 被引用 10 次
- Gradient-based Discrete Sampling with Automatic Cyclical SchedulingPatrick Pynadath, Riddhiman Bhattacharya, Arun Hariharan, Ruqi ZhangNeurIPS 2024 · 被引用 10 次
