Deep MMD Gradient Flow without adversarial training
Alexandre Galashov, Valentin De Bortoli, Arthur Gretton
Abstract
We propose a gradient flow procedure for generative modeling by transporting particles from an initial source distribution to a target distribution, where the gradient field on the particles is given by a noise-adaptive Wasserstein Gradient of the Maximum Mean Discrepancy (MMD). The noise-adaptive MMD is trained on data distributions corrupted by increasing levels of noise, obtained via a forward diffusion process, as commonly used in denoising diffusion probabilistic models. The result is a generalization of MMD Gradient Flow, which we call Diffusion-MMD-Gradient Flow or DMMD. The divergence training procedure is related to discriminator training in Generative Adversarial Networks (GAN), but does not require adversarial training. We obtain competitive empirical performance in unconditional image generation on CIFAR10, MNIST, CELEB-A (64 x64) and LSUN Church (64 x 64). Furthermore, we demonstrate the validity of the approach when MMD is replaced by a lower bound on the KL divergence.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext fa2976c8-0215-4fa6-8a3c-69c630b63f0fCited by top-tier papers8
- Scale-wise Distillation of Diffusion ModelsNikita Starodubcev, Ilya Drobyshevskiy, Denis Kuznedelev, Artem Babenko et al.ICLR 2026 · 13 citations
- Statistical and Geometrical properties of the Kernel Kullback-Leibler divergenceAnna Korba, Francis R. Bach, Clémentine ChazalNeurIPS 2024 · 5 citations
- Doubly-Robust Estimation of Counterfactual Policy Mean EmbeddingsHoussam Zenati, Bariscan Bozkurt, Arthur GrettonNeurIPS 2025 · 3 citations
- Sequence Modeling with Spectral Mean FlowsJinwoo Kim, Max Beier, Petar Bevanda, Nayun Kim et al.NeurIPS 2025 · 2 citations
- Distributional Diffusion Models with Scoring RulesValentin De Bortoli, Alexandre Galashov, J. Swaroop Guntupalli, Guangyao Zhou et al.ICML 2025
Builds on22
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li et al.NeurIPS 2022 · 8,965 citations
- Training Generative Adversarial Networks with Limited DataTero Karras, Miika Aittala, Janne Hellsten, Samuli Laine et al.NeurIPS 2020 · 2,345 citations
- Consistency ModelsYang Song, Prafulla Dhariwal, Mark Chen, Ilya SutskeverICML 2023 · 1,720 citations
Related papers
- MMD Guidance: Training-Free Distribution Adaptation for Diffusion Models via Maximum Mean Discrepancy GuidanceMatina Mahdizadeh Sani, Nima Jamali, Mohammad Jalali, Farzan FarniaICML 2026 · 4 citations
- MonoFlow: Rethinking Divergence GANs via the Perspective of Wasserstein Gradient FlowsMingxuan Yi, Zhanxing Zhu, Song LiuICML 2023 · 18 citations
- Posterior Sampling Based on Gradient Flows of the MMD with Negative Distance KernelPaul Hagemann, Johannes Hertrich, Fabian Altekrüger, Robert Beinert et al.ICLR 2024 · 32 citations
- Gradual Domain Adaptation via Gradient FlowZhan Zhuang, Yu Zhang, Ying WeiICLR 2024 · 15 citations
- A Characteristic Function Approach to Deep Implicit Generative ModelingAbdul Fatir Ansari, Jonathan Scarlett, Harold SohCVPR 2020
