Glauber Generative Model: Discrete Diffusion Models via Binary Classification
Harshit Varma, Dheeraj Mysore Nagaraj, Karthikeyan Shanmugam
摘要
We introduce the Glauber Generative Model (GGM), a new class of discrete diffusion models, to obtain new samples from a distribution given samples from a discrete space. GGM deploys a discrete Markov chain called the heat bath dynamics (or the Glauber dynamics) to denoise a sequence of noisy tokens to a sample from a joint distribution of discrete tokens. Our novel conceptual framework provides an exact reduction of the task of learning the denoising Markov chain to solving a class of binary classification tasks. More specifically, the model learns to classify a given token in a noisy sequence as signal or noise. In contrast, prior works on discrete diffusion models either solve regression problems to learn importance ratios, or minimize loss functions given by variational approximations. We apply GGM to language modeling and image generation, where images are discretized using image tokenizers like VQGANs. We show that it outperforms existing discrete diffusion models in language generation, and demonstrates strong performance for image generation without using dataset-specific image tokenizers. We also show that our model is capable of performing well in zero-shot control settings like text and image infilling. Πt(a) Πt(ϕ) 1 ŷa -1 .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Fast Solvers for Discrete Diffusion Models: Theory and Applications of High-Order AlgorithmsYinuo Ren, Haoxuan Chen, Yuchen Zhu, Wei Guo 等NeurIPS 2025 · 被引用 51 次
- Fine-Tuning Masked Diffusion for Provable Self-CorrectionJaeyeon Kim, Seunggeun Kim, Taekyun Lee, David Pan 等ICML 2026 · 被引用 35 次
- Anchored Diffusion Language ModelLitu Rout, Constantine Caramanis, Sanjay ShakkottaiNeurIPS 2025 · 被引用 19 次
- Dimension-free Score Matching and Time Bootstrapping for Diffusion ModelsSyamantak Kumar, Dheeraj Nagaraj, Purnamrita SarkarNeurIPS 2025 · 被引用 2 次
- Train for the Worst, Plan for the Best: Understanding Token Ordering in Masked DiffusionsJaeyeon Kim, Kulin Shah, Vasilis Kontonis, Sham M. Kakade 等ICML 2025
它引用的顶会 Paper28
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Zero-Shot Text-to-Image GenerationAditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray 等ICML 2021 · 被引用 6,356 次
- Scalable Diffusion Models with TransformersWilliam Peebles, Saining XieICCV 2023 · 被引用 5,568 次
相关 Paper
- Autoregressive Image Generation without Vector QuantizationTianhong Li, Yonglong Tian, He Li, Mingyang Deng 等NeurIPS 2024 · 被引用 758 次
- Discrete Modeling via Boundary Conditional Diffusion ProcessesYuxuan Gu, Xiaocheng Feng, Lei Huang, Yingsheng Wu 等NeurIPS 2024
- Simple Guidance Mechanisms for Discrete Diffusion ModelsYair Schiff, Subham Sekhar Sahoo, Hao Phung, Guanghan Wang 等ICLR 2025
- Optimality of FSQ Tokens for Continuous Diffusion for Categorical Data with Application to Text-to-SpeechVadim Popov, Wenju Gu, Tasnima Sadekova, Georgii Aparin 等ICML 2026
- A Cheaper and Better Diffusion Language Model with Soft-Masked NoiseJiaao Chen, Aston Zhang, Mu Li, Alex Smola 等EMNLP 2023 · 被引用 16 次
