Gradient-based Discrete Sampling with Automatic Cyclical Scheduling
Patrick Pynadath, Riddhiman Bhattacharya, Arun Hariharan, Ruqi Zhang
摘要
Discrete distributions, particularly in high-dimensional deep models, are often highly multimodal due to inherent discontinuities. While gradient-based discrete sampling has proven effective, it is susceptible to becoming trapped in local modes due to the gradient information. To tackle this challenge, we propose an automatic cyclical scheduling, designed for efficient and accurate sampling in multimodal discrete distributions. Our method contains three key components: (1) a cyclical step size schedule where large steps discover new modes and small steps exploit each mode; (2) a cyclical balancing schedule, ensuring"balanced"proposals for given step sizes and high efficiency of the Markov chain; and (3) an automatic tuning scheme for adjusting the hyperparameters in the cyclical schedules, allowing adaptability across diverse datasets with minimal tuning. We prove the non-asymptotic convergence and inference guarantee for our method in general discrete distributions. Extensive experiments demonstrate the superiority of our method in sampling complex multimodal discrete distributions.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Exploring Non-Convex Discrete Energy Landscapes: An Efficient Langevin-Like Sampler with Replica ExchangeHaoyang Zheng, Hengrong Du, Ruqi Zhang, Guang LinAAAI 2026
- Beyond Self-Repellent Kernels: History-Driven Target Towards Efficient Nonlinear MCMC on General GraphsJie Hu, Yi-Ting Ma, Do Young EunICML 2025
- Controlled LLM Decoding via Discrete Auto-regressive BiasingPatrick Pynadath, Ruqi ZhangICLR 2025
- From Predictors to Samplers via the Training TrajectorySoumya Ram, Akhila RamICLR 2026
它引用的顶会 Paper10
- Cyclical Stochastic Gradient MCMC for Bayesian Deep LearningRuqi Zhang, Chunyuan Li, Jianyi Zhang, Changyou Chen 等ICLR 2020 · 被引用 292 次
- Oops I Took A Gradient: Scalable Sampling for Discrete DistributionsWill Grathwohl, Kevin Swersky, Milad Hashemi, David Duvenaud 等ICML 2021 · 被引用 113 次
- Non-convex Learning via Replica Exchange Stochastic Gradient MCMCWei Deng, Qi Feng, Liyao Gao, Faming Liang 等ICML 2020 · 被引用 54 次
- A Langevin-like Sampler for Discrete DistributionsRuqi Zhang, Xingchao Liu, Qiang LiuICML 2022 · 被引用 51 次
- A Contour Stochastic Gradient Langevin Dynamics Algorithm for Simulations of Multi-modal DistributionsWei Deng, Guang Lin, Faming LiangNeurIPS 2020 · 被引用 37 次
相关 Paper
- Learning to Explore for Stochastic Gradient MCMCSeunghyun Kim, Seohyeon Jung, Seonghyeon Kim, Juho LeeICML 2024 · 被引用 2 次
- AutoSampling: Search for Effective Data Sampling SchedulesMing Sun, Haoxuan Dou, Baopu Li, Junjie Yan 等ICML 2021 · 被引用 6 次
- Optimal Scaling for Locally Balanced Proposals in Discrete SpacesHaoran Sun, Hanjun Dai, Dale SchuurmansNeurIPS 2022 · 被引用 14 次
- Entropic Mirror Monte CarloAnas CHERRADI, Yazid Janati, Alain Oliviero Durmus, Sylvain Le Corff 等ICML 2026
- Any-scale Balanced Samplers for Discrete SpaceHaoran Sun, Bo Dai, Charles Sutton, Dale Schuurmans 等ICLR 2023
