First Hitting Diffusion Models for Generating Manifold, Graph and Categorical Data
Mao Ye, Lemeng Wu, Qiang Liu
Abstract
We propose a family of First Hitting Diffusion Models (FHDM), deep generative models that generate data with a diffusion process that terminates at a random first hitting time. This yields an extension of the standard fixed-time diffusion models that terminate at a pre-specified deterministic time. Although standard diffusion models are designed for continuous unconstrained data, FHDM is naturally designed to learn distributions on continuous as well as a range of discrete and structure domains. Moreover, FHDM enables instance-dependent terminate time and accelerates the diffusion process to sample higher quality data with fewer diffusion steps. Technically, we train FHDM by maximum likelihood estimation on diffusion trajectories augmented from observed data with conditional first hitting processes (i.e., bridge) derived based on Doob's -transform, deviating from the commonly used time-reversal mechanism. We apply FHDM to generate data in various domains such as point cloud (general continuous distribution), climate and geographical events on earth (continuous distribution on the sphere), unweighted graphs (distribution of binary matrices), and segmentation maps of 2D images (high-dimensional categorical distribution). We observe considerable improvement compared with the state-of-the-art approaches in both quality and speed.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers12
- InstaFlow: One Step is Enough for High-Quality Diffusion-Based Text-to-Image GenerationXingchao Liu, Xiwen Zhang, Jianzhu Ma, Jian Peng et al.ICLR 2024 · 358 citations
- Mirror Diffusion Models for Constrained and Watermarked GenerationGuan-Horng Liu, Tianrong Chen, Evangelos A. Theodorou, Molei TaoNeurIPS 2023 · 55 citations
- DEFT: Efficient Fine-tuning of Diffusion Models by Learning the Generalised -transformAlexander Denker, Francisco Vargas, Shreyas Padhy, Kieran Didi et al.NeurIPS 2024 · 52 citations
- Trajectory Diffusion for ObjectGoal NavigationXinyao Yu, Sixian Zhang, Xinhang Song, Xiaorong Qin et al.NeurIPS 2024 · 32 citations
- Blackout Diffusion: Generative Diffusion Models in Discrete-State SpacesJavier E. Santos, Zachary R. Fox, Nicholas Lubbers, Yen Ting LinICML 2023 · 28 citations
Builds on14
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
- DiffWave: A Versatile Diffusion Model for Audio SynthesisZhifeng Kong, Wei Ping, Jiaji Huang, Kexin Zhao et al.ICLR 2021 · 1,902 citations
- Improved Techniques for Training Score-Based Generative ModelsYang Song, Stefano ErmonNeurIPS 2020 · 1,527 citations
- Score-Based Generative Modeling through Stochastic Differential EquationsYang Song, Jascha Sohl-Dickstein, Diederik P. Kingma, Abhishek Kumar et al.ICLR 2021 · 1,270 citations
Related papers
- Masked Diffusion Models are Secretly Time-Agnostic Masked Models and Exploit Inaccurate Categorical SamplingKaiwen Zheng, Yongxin Chen, Hanzi Mao, Ming-Yu Liu et al.ICLR 2025
- Learning Diffusion Bridges on Constrained DomainsXingchao Liu, Lemeng Wu, Mao Ye, Qiang LiuICLR 2023
- A Continuous Time Framework for Discrete Denoising ModelsAndrew Campbell, Joe Benton, Valentin De Bortoli, Thomas Rainforth et al.NeurIPS 2022 · 496 citations
- Learning to Jump: Thinning and Thickening Latent Counts for Generative ModelingTianqi Chen, Mingyuan ZhouICML 2023 · 13 citations
- Discrete-state Continuous-time Diffusion for Graph GenerationZhe Xu, Ruizhong Qiu, Yuzhong Chen, Huiyuan Chen et al.NeurIPS 2024 · 92 citations
