Can Push-forward Generative Models Fit Multimodal Distributions?
Antoine Salmona, Valentin De Bortoli, Julie Delon, Agnès Desolneux
Abstract
Many generative models synthesize data by transforming a standard Gaussian random variable using a deterministic neural network. Among these models are the Variational Autoencoders and the Generative Adversarial Networks. In this work, we call them"push-forward"models and study their expressivity. We show that the Lipschitz constant of these generative networks has to be large in order to fit multimodal distributions. More precisely, we show that the total variation distance and the Kullback-Leibler divergence between the generated and the data distribution are bounded from below by a constant depending on the mode separation and the Lipschitz constant. Since constraining the Lipschitz constants of neural networks is a common way to stabilize generative models, there is a provable trade-off between the ability of push-forward models to approximate multimodal distributions and the stability of their training. We validate our findings on one-dimensional and image datasets and empirically show that generative models consisting of stacked networks with stochastic input at each step, such as diffusion models do not suffer of such limitations.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext bcfddbb2-1afd-4656-8e17-9c11c44c632bCited by top-tier papers17
- Unpaired Image-to-Image Translation via Neural Schrödinger BridgeBeomsu Kim, Gihyun Kwon, Kwanyoung Kim, Jong Chul YeICLR 2024 · 131 citations
- Metric Flow Matching for Smooth Interpolations on the Data ManifoldKacper Kapusniak, Peter Potaptchik, Teodora Reu, Leo Zhang et al.NeurIPS 2024 · 89 citations
- Fast ODE-based Sampling for Diffusion Models in Around 5 StepsZhenyu Zhou, Defang Chen, Can Wang, Chun ChenCVPR 2024 · 22 citations
- Particle Semi-Implicit Variational InferenceJen Ning Lim, Adam M. JohansenNeurIPS 2024 · 13 citations
- Generative Learning for Solving Non-Convex Problem with Multi-Valued Input-Solution MappingEnming Liang, Minghua ChenICLR 2024 · 10 citations
Builds on10
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
- Score-Based Generative Modeling through Stochastic Differential EquationsYang Song, Jascha Sohl-Dickstein, Diederik P. Kingma, Abhishek Kumar et al.ICLR 2021 · 1,270 citations
- The Intrinsic Dimension of Images and Its Impact on LearningPhillip Pope, Chen Zhu, Ahmed Abdelkader, Micah Goldblum et al.ICLR 2021 · 381 citations
- Relaxing Bijectivity Constraints with Continuously Indexed Normalising FlowsRobert Cornish, Anthony L. Caterini, George Deligiannidis, Arnaud DoucetICML 2020 · 141 citations
Related papers
- Unraveling the Smoothness Properties of Diffusion Models: A Gaussian Mixture PerspectiveYingyu Liang, Zhizhou Sha, Zhenmei Shi, Zhao Song et al.ICCV 2025 · 23 citations
- Shape your Space: A Gaussian Mixture Regularization Approach to Deterministic AutoencodersAmrutha Saseendran, Kathrin Skubch, Stefan Falkner, Margret KeuperNeurIPS 2021 · 13 citations
- Diffusion Normalizing FlowQinsheng Zhang, Yongxin ChenNeurIPS 2021 · 119 citations
- The Effects of Invertibility on the Representational Complexity of Encoders in Variational AutoencodersDivyansh Pareek, Andrej RisteskiICLR 2022
- Statistically Optimal Generative Modeling with Maximum Deviation from the Empirical DistributionElen Vardanyan, Sona Hunanyan, Tigran Galstyan, Arshak Minasyan et al.ICML 2024 · 3 citations
