Learning to Jump: Thinning and Thickening Latent Counts for Generative Modeling
Tianqi Chen, Mingyuan Zhou
摘要
Learning to denoise has emerged as a prominent paradigm to design state-of-the-art deep generative models for natural images. How to use it to model the distributions of both continuous real-valued data and categorical data has been well studied in recently proposed diffusion models. However, it is found in this paper to have limited ability in modeling some other types of data, such as count and non-negative continuous data, that are often highly sparse, skewed, heavy-tailed, and/or overdispersed. To this end, we propose learning to jump as a general recipe for generative modeling of various types of data. Using a forward count thinning process to construct learning objectives to train a deep neural network, it employs a reverse count thickening process to iteratively refine its generation through that network. We demonstrate when learning to jump is expected to perform comparably to learning to denoise, and when it is expected to perform better. For example, learning to jump is recommended when the training data is non-negative and exhibits strong sparsity, skewness, heavy-tailedness, and/or heterogeneity.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- Add and Thin: Diffusion for Temporal Point ProcessesDavid Lüdke, Marin Bilos, Oleksandr Shchur, Marten Lienen 等NeurIPS 2023 · 被引用 34 次
- FS-DFM: Fast and Accurate Long Text Generation with Few-Step Diffusion Language ModelsAmin Karimi Monsefi, Nikhil Bhendawade, Manuel Rafael Ciosici, Dominic Culver 等ICLR 2026 · 被引用 15 次
- ItDPDM: Information-Theoretic Discrete Poisson Diffusion ModelSagnik Bhattacharya, Abhiram Rao Gorle, Ahsan Bilal, Connor Ding 等NeurIPS 2025 · 被引用 6 次
- Score Forgetting Distillation: A Swift, Data-Free Method for Machine Unlearning in Diffusion ModelsTianqi Chen, Shujian Zhang, Mingyuan ZhouICLR 2025 · 被引用 2 次
- Guided Score identity Distillation for Data-Free One-Step Text-to-Image GenerationMingyuan Zhou, Zhendong Wang, Huangjie Zheng, Hai HuangICLR 2025
它引用的顶会 Paper36
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li 等NeurIPS 2022 · 被引用 8,965 次
相关 Paper
- Generalization of Diffusion Models Arises with a Balanced Representation SpaceZekai Zhang, Xiao Li, Xiang Li, Lianghe Shi 等ICLR 2026 · 被引用 14 次
- A Continuous Time Framework for Discrete Denoising ModelsAndrew Campbell, Joe Benton, Valentin De Bortoli, Thomas Rainforth 等NeurIPS 2022 · 被引用 496 次
- Score-based Continuous-time Discrete Diffusion ModelsHaoran Sun, Lijun Yu, Bo Dai, Dale Schuurmans 等ICLR 2023 · 被引用 7 次
- Argmax Flows and Multinomial Diffusion: Learning Categorical DistributionsEmiel Hoogeboom, Didrik Nielsen, Priyank Jaini, Patrick Forré 等NeurIPS 2021 · 被引用 782 次
- First Hitting Diffusion Models for Generating Manifold, Graph and Categorical DataMao Ye, Lemeng Wu, Qiang LiuNeurIPS 2022 · 被引用 25 次
