What Is It Like to Be a Noise? An Entropy-based Gaussian Noise Regularization for Diffusion Models
Pascal Chang, Kai Lascheit, Jingwei Tang, Markus Gross, Vinicius C. Azevedo
摘要
Inference-time optimization of diffusion latents enables powerful control but often degrades the statistical structure of true Gaussian noise, causing artifacts and reward hacking. To address this, we propose a Gaussianity regularizer that aligns a sample's local statistics with a typical Gaussian realization, rather than relying on pointwise likelihood. We formalize this by computing the KL divergence between the sample distribution and the Gaussian prior. To make the divergence computation tractable from a single sample, we lift each candidate latent into an empirical distribution induced by its statistics and model it as a pairwise Markov Random Field. This yields a Bethe-Kikuchi-style regularizer with 1D marginal, 2D spatial, and multi-scale terms. Our results show improved latent optimization stability and generation quality over prior approaches.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper36
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Training Diffusion Models with Reinforcement LearningKevin Black, Michael Janner, Yilun Du, Ilya Kostrikov 等ICLR 2024 · 被引用 816 次
- Zero-shot Image-to-Image TranslationGaurav Parmar, Krishna Kumar Singh, Richard Zhang, Yijun Li 等SIGGRAPH 2023 · 被引用 355 次
- Preserve Your Own Correlation: A Noise Prior for Video Diffusion ModelsSongwei Ge, Seungjun Nah, Guilin Liu, Tyler Poon 等ICCV 2023 · 被引用 319 次
- FreeNoise: Tuning-Free Longer Video Diffusion via Noise ReschedulingHaonan Qiu, Menghan Xia, Yong Zhang, Yingqing He 等ICLR 2024 · 被引用 171 次
相关 Paper
- Moment- and Power-Spectrum-Based Gaussianity Regularization for Text-to-Image ModelsJisung Hwang, Jaihoon Kim, Minhyuk SungNeurIPS 2025 · 被引用 2 次
- DeRaDiff: Denoising Time Realignment of Diffusion ModelsRatnavibusena Don Shahain Manujith, Teoh Tze Tzun, Kenji Kawaguchi, Yang ZhangICLR 2026 · 被引用 1 次
- Information Theoretic Learning for Diffusion Models with Warm StartYirong Shen, Lu Gan, Cong LingNeurIPS 2025 · 被引用 4 次
- The Universal Normal EmbeddingChen Tasker, Roy Betser, Eyal Gofer, Meir Yossef Levi 等CVPR 2026 · 被引用 4 次
- Isometric Representation Learning for Disentangled Latent Space of Diffusion ModelsJaehoon Hahm, Junho Lee, Sunghyun Kim, Joonseok LeeICML 2024 · 被引用 21 次
