A Langevin-like Sampler for Discrete Distributions
Ruqi Zhang, Xingchao Liu, Qiang Liu
Abstract
We propose discrete Langevin proposal (DLP), a simple and scalable gradient-based proposal for sampling complex high-dimensional discrete distributions. In contrast to Gibbs sampling-based methods, DLP is able to update all coordinates in parallel in a single step and the magnitude of changes is controlled by a stepsize. This allows a cheap and efficient exploration in the space of high-dimensional and strongly correlated variables. We prove the efficiency of DLP by showing that the asymptotic bias of its stationary distribution is zero for log-quadratic distributions, and is small for distributions that are close to being log-quadratic. With DLP, we develop several variants of sampling algorithms, including unadjusted, Metropolis-adjusted, stochastic and preconditioned versions. DLP outperforms many popular alternatives on a wide variety of tasks, including Ising models, restricted Boltzmann machines, deep energy-based models, binary neural networks and language generation.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers26
- Automatically Auditing Large Language Models via Discrete OptimizationErik Jones, Anca D. Dragan, Aditi Raghunathan, Jacob SteinhardtICML 2023 · 232 citations
- Generative Flow Networks for Discrete Probabilistic ModelingDinghuai Zhang, Nikolay Malkin, Zhen Liu, Alexandra Volokhova et al.ICML 2022 · 131 citations
- BayesDAG: Gradient-Based Posterior Inference for Causal DiscoveryYashas Annadani, Nick Pawlowski, Joel Jennings, Stefan Bauer et al.NeurIPS 2023 · 54 citations
- Revisiting Sampling for Combinatorial OptimizationHaoran Sun, Katayoon Goshvadi, Azade Nova, Dale Schuurmans et al.ICML 2023 · 28 citations
- Alleviating Adversarial Attacks on Variational Autoencoders with MCMCAnna Kuzina, Max Welling, Jakub M. TomczakNeurIPS 2022 · 16 citations
Builds on9
- Your classifier is secretly an energy based model and you should treat it like oneWill Grathwohl, Kuan-Chieh Wang, Jörn-Henrik Jacobsen, David Duvenaud et al.ICLR 2020 · 643 citations
- Cyclical Stochastic Gradient MCMC for Bayesian Deep LearningRuqi Zhang, Chunyuan Li, Jianyi Zhang, Changyou Chen et al.ICLR 2020 · 292 citations
- Understanding Knowledge Distillation in Non-autoregressive Machine TranslationChunting Zhou, Jiatao Gu, Graham NeubigICLR 2020 · 235 citations
- Generative Flow Networks for Discrete Probabilistic ModelingDinghuai Zhang, Nikolay Malkin, Zhen Liu, Alexandra Volokhova et al.ICML 2022 · 131 citations
- Oops I Took A Gradient: Scalable Sampling for Discrete DistributionsWill Grathwohl, Kevin Swersky, Milad Hashemi, David Duvenaud et al.ICML 2021 · 113 citations
Related papers
- Optimal Scaling for Locally Balanced Proposals in Discrete SpacesHaoran Sun, Hanjun Dai, Dale SchuurmansNeurIPS 2022 · 14 citations
- Exploring Non-Convex Discrete Energy Landscapes: An Efficient Langevin-Like Sampler with Replica ExchangeHaoyang Zheng, Hengrong Du, Ruqi Zhang, Guang LinAAAI 2026
- Optimal Underdamped Langevin MCMC MethodZhengmian Hu, Feihu Huang, Heng HuangNeurIPS 2021 · 5 citations
- LSB: Local Self-Balancing MCMC in Discrete SpacesEmanuele SansoneICML 2022 · 10 citations
- Gradient-based Discrete Sampling with Automatic Cyclical SchedulingPatrick Pynadath, Riddhiman Bhattacharya, Arun Hariharan, Ruqi ZhangNeurIPS 2024 · 10 citations
