Gradient-based Discrete Sampling with Automatic Cyclical Scheduling
Patrick Pynadath, Riddhiman Bhattacharya, Arun Hariharan, Ruqi Zhang
Abstract
Discrete distributions, particularly in high-dimensional deep models, are often highly multimodal due to inherent discontinuities. While gradient-based discrete sampling has proven effective, it is susceptible to becoming trapped in local modes due to the gradient information. To tackle this challenge, we propose an automatic cyclical scheduling, designed for efficient and accurate sampling in multimodal discrete distributions. Our method contains three key components: (1) a cyclical step size schedule where large steps discover new modes and small steps exploit each mode; (2) a cyclical balancing schedule, ensuring"balanced"proposals for given step sizes and high efficiency of the Markov chain; and (3) an automatic tuning scheme for adjusting the hyperparameters in the cyclical schedules, allowing adaptability across diverse datasets with minimal tuning. We prove the non-asymptotic convergence and inference guarantee for our method in general discrete distributions. Extensive experiments demonstrate the superiority of our method in sampling complex multimodal discrete distributions.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f17e1556-5f03-4036-bb25-0afab5c80924Cited by top-tier papers4
- Exploring Non-Convex Discrete Energy Landscapes: An Efficient Langevin-Like Sampler with Replica ExchangeHaoyang Zheng, Hengrong Du, Ruqi Zhang, Guang LinAAAI 2026
- Beyond Self-Repellent Kernels: History-Driven Target Towards Efficient Nonlinear MCMC on General GraphsJie Hu, Yi-Ting Ma, Do Young EunICML 2025
- Controlled LLM Decoding via Discrete Auto-regressive BiasingPatrick Pynadath, Ruqi ZhangICLR 2025
- From Predictors to Samplers via the Training TrajectorySoumya Ram, Akhila RamICLR 2026
Builds on10
- Cyclical Stochastic Gradient MCMC for Bayesian Deep LearningRuqi Zhang, Chunyuan Li, Jianyi Zhang, Changyou Chen et al.ICLR 2020 · 292 citations
- Oops I Took A Gradient: Scalable Sampling for Discrete DistributionsWill Grathwohl, Kevin Swersky, Milad Hashemi, David Duvenaud et al.ICML 2021 · 113 citations
- Non-convex Learning via Replica Exchange Stochastic Gradient MCMCWei Deng, Qi Feng, Liyao Gao, Faming Liang et al.ICML 2020 · 54 citations
- A Langevin-like Sampler for Discrete DistributionsRuqi Zhang, Xingchao Liu, Qiang LiuICML 2022 · 51 citations
- A Contour Stochastic Gradient Langevin Dynamics Algorithm for Simulations of Multi-modal DistributionsWei Deng, Guang Lin, Faming LiangNeurIPS 2020 · 37 citations
Related papers
- Learning to Explore for Stochastic Gradient MCMCSeunghyun Kim, Seohyeon Jung, Seonghyeon Kim, Juho LeeICML 2024 · 2 citations
- AutoSampling: Search for Effective Data Sampling SchedulesMing Sun, Haoxuan Dou, Baopu Li, Junjie Yan et al.ICML 2021 · 6 citations
- Optimal Scaling for Locally Balanced Proposals in Discrete SpacesHaoran Sun, Hanjun Dai, Dale SchuurmansNeurIPS 2022 · 14 citations
- Entropic Mirror Monte CarloAnas CHERRADI, Yazid Janati, Alain Oliviero Durmus, Sylvain Le Corff et al.ICML 2026
- Any-scale Balanced Samplers for Discrete SpaceHaoran Sun, Bo Dai, Charles Sutton, Dale Schuurmans et al.ICLR 2023
