Learning multi-scale local conditional probability models of images
Zahra Kadkhodaie, Florentin Guth, Stéphane Mallat, Eero P. Simoncelli
摘要
Deep neural networks can learn powerful prior probability models for images, as evidenced by the high-quality generations obtained with recent score-based diffusion methods. But the means by which these networks capture complex global statistical structure, apparently without suffering from the curse of dimensionality, remain a mystery. To study this, we incorporate diffusion methods into a multi-scale decomposition, reducing dimensionality by assuming a stationary local Markov model for wavelet coefficients conditioned on coarser-scale coefficients. We instantiate this model using convolutional neural networks (CNNs) with local receptive fields, which enforce both the stationarity and Markov properties. Global structures are captured using a CNN with receptive fields covering the entire (but small) low-pass image. We test this model on a dataset of face images, which are highly non-stationary and contain large-scale geometric structures. Remarkably, denoising, super-resolution, and image synthesis results all demonstrate that these structures can be captured with significantly smaller conditioning neighborhoods than required by a Markov model implemented in the pixel domain. Our results show that score estimation for large complex images can be reduced to low-dimensional Markov conditional models across scales, alleviating the curse of dimensionality. Deep neural networks (DNNs) have produced dramatic advances in synthesizing complex images and solving inverse problems, all of which rely (at least implicitly) on prior probability models. Of particular note is the recent development of "diffusion methods" (Sohl-Dickstein et al., 2015), in which a network trained for image denoising is incorporated into an iterative algorithm to draw samples from the prior (
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- Generalization in diffusion models arises from geometry-adaptive harmonic representationsZahra Kadkhodaie, Florentin Guth, Eero P. Simoncelli, Stéphane MallatICLR 2024 · 被引用 168 次
- Matryoshka Diffusion ModelsJiatao Gu, Shuangfei Zhai, Yizhe Zhang, Joshua Susskind 等ICLR 2024 · 被引用 73 次
- UDPM: Upsampling Diffusion Probabilistic ModelsShady Abu-Hussein, Raja GiryesNeurIPS 2024 · 被引用 10 次
- Conditionally Strongly Log-Concave Generative ModelsFlorentin Guth, Etienne Lempereur, Joan Bruna, Stéphane MallatICML 2023 · 被引用 5 次
- Diffusion Transformer Meets Multi-Level Wavelet Spectrum for Single Image Super-ResolutionPeng Du, Hui Li, Han Xu, Paul Barom Jeon 等ICCV 2025 · 被引用 3 次
它引用的顶会 Paper7
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- SNIPS: Solving Noisy Inverse Problems StochasticallyBahjat Kawar, Gregory Vaksman, Michael EladNeurIPS 2021 · 被引用 263 次
- Stochastic Solutions for Linear Inverse Problems using the Prior Implicit in a DenoiserZahra Kadkhodaie, Eero P. SimoncelliNeurIPS 2021 · 被引用 202 次
- Robust And Interpretable Blind Image Denoising Via Bias-Free Convolutional Neural NetworksSreyas Mohan, Zahra Kadkhodaie, Eero P. Simoncelli, Carlos Fernandez-GrandaICLR 2020 · 被引用 154 次
- Wavelet Score-Based Generative ModelingFlorentin Guth, Simon Coste, Valentin De Bortoli, Stéphane MallatNeurIPS 2022 · 被引用 98 次
相关 Paper
- Score Approximation, Estimation and Distribution Recovery of Diffusion Models on Low-Dimensional DataMinshuo Chen, Kaixuan Huang, Tuo Zhao, Mengdi WangICML 2023 · 被引用 168 次
- A Restoration Network as an Implicit PriorYuyang Hu, Mauricio Delbracio, Peyman Milanfar, Ulugbek KamilovICLR 2024 · 被引用 17 次
- Locality in Image Diffusion Models Emerges from Data StatisticsArtem Lukoianov, Chenyang Yuan, Justin M. Solomon, Vincent SitzmannNeurIPS 2025 · 被引用 32 次
- Learning to Deblur Face Images via Sketch SynthesisSongnan Lin, Jiawei Zhang, Jinshan Pan, Yicun Liu 等AAAI 2020 · 被引用 26 次
- Deep Gaussian Markov Random FieldsPer Sidén, Fredrik LindstenICML 2020 · 被引用 25 次
