Lossy Image Compression with Conditional Diffusion Models
Ruihan Yang, Stephan Mandt
摘要
This paper outlines an end-to-end optimized lossy image compression framework using diffusion generative models. The approach relies on the transform coding paradigm, where an image is mapped into a latent space for entropy coding and, from there, mapped back to the data space for reconstruction. In contrast to VAE-based neural compression, where the (mean) decoder is a deterministic neural network, our decoder is a conditional diffusion model. Our approach thus introduces an additional content'' latent variable on which the reverse diffusion process is conditioned and uses this variable to store information about the image. The remaining texture'' variables characterizing the diffusion process are synthesized at decoding time. We show that the model's performance can be tuned toward perceptual metrics of interest. Our extensive experiments involving multiple datasets and image quality assessment metrics show that our approach yields stronger reported FID scores than the GAN-based model, while also yielding competitive performance with VAE-based models in several distortion metrics. Furthermore, training the diffusion with -parameterization enables high-quality reconstructions in only a handful of decoding steps, greatly affecting the model's practicality. Our code is available at: https://github.com/buggyyang/CDC_compression
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper62
- Towards image compression with perfect realism at ultra-low bitratesMarlène Careil, Matthew J. Muckley, Jakob Verbeek, Stéphane LathuilièreICLR 2024 · 被引用 122 次
- Improving Statistical Fidelity for Neural Image Compression with Implicit Local Likelihood ModelsMatthew J. Muckley, Alaaeldin El-Nouby, Karen Ullrich, Hervé Jégou 等ICML 2023 · 被引用 114 次
- Diffusion Models With Learned Adaptive NoiseSubham S. Sahoo, Aaron Gokaslan, Christopher De Sa, Volodymyr KuleshovNeurIPS 2024 · 被引用 64 次
- Improved Techniques for Maximum Likelihood Estimation for Diffusion ODEsKaiwen Zheng, Cheng Lu, Jianfei Chen, Jun ZhuICML 2023 · 被引用 55 次
- Neural Flow Diffusion Models: Learnable Forward Process for Improved Diffusion ModellingGrigory Bartosh, Dmitry P. Vetrov, Christian Andersson NaessethNeurIPS 2024 · 被引用 49 次
它引用的顶会 Paper21
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
- Video Diffusion ModelsJonathan Ho, Tim Salimans, Alexey A. Gritsenko, William Chan 等NeurIPS 2022 · 被引用 2,948 次
- Score-Based Generative Modeling through Stochastic Differential EquationsYang Song, Jascha Sohl-Dickstein, Diederik P. Kingma, Abhishek Kumar 等ICLR 2021 · 被引用 1,270 次
相关 Paper
- Correcting Diffusion-Based Perceptual Image Compression with Privileged End-to-End DecoderYiyang Ma, Wenhan Yang, Jiaying LiuICML 2024 · 被引用 12 次
- Compressed Image Generation with Denoising Diffusion Codebook ModelsGuy Ohayon, Hila Manor, Tomer Michaeli, Michael EladICML 2025
- Laplacian-guided Entropy Model in Neural Codec with Blur-dissipated SynthesisAtefeh Khoshkhahtinat, Ali Zafari, Piyush M. Mehta, Nasser M. NasrabadiCVPR 2024
- CoD: A Diffusion Foundation Model for Image CompressionZhaoyang Jia, Zihan Zheng, Naifu Xue, Jiahao Li 等CVPR 2026 · 被引用 9 次
- Progressive Compression with Universally Quantized Diffusion ModelsYibo Yang, Justus C. Will, Stephan MandtICLR 2025
