Multi-Realism Image Compression with a Conditional Generator
Eirikur Agustsson, David Minnen, George Toderici, Fabian Mentzer
摘要
By optimizing the rate-distortion-realism trade-off, generative compression approaches produce detailed, realistic images, even at low bit rates, instead of the blurry reconstructions produced by rate-distortion optimized models. However, previous methods do not explicitly control how much detail is synthesized, which results in a common criticism of these methods: users might be worried that a misleading reconstruction far from the input image is generated. In this work, we alleviate these concerns by training a decoder that can bridge the two regimes and navigate the distortion-realism trade-off. From a single compressed representation, the receiver can decide to either reconstruct a low mean squared error reconstruction that is close to the input, a realistic reconstruction with high perceptual quality, or anything in between. With our method, we set a new state-of-the-art in distortion-realism, pushing the frontier of achievable distortion-realism pairs, i.e., our method achieves better distortions at high realism and better realism at low distortion than ever before.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper37
- Lossy Image Compression with Conditional Diffusion ModelsRuihan Yang, Stephan MandtNeurIPS 2023 · 被引用 268 次
- Towards image compression with perfect realism at ultra-low bitratesMarlène Careil, Matthew J. Muckley, Jakob Verbeek, Stéphane LathuilièreICLR 2024 · 被引用 122 次
- Improving Statistical Fidelity for Neural Image Compression with Implicit Local Likelihood ModelsMatthew J. Muckley, Alaaeldin El-Nouby, Karen Ullrich, Hervé Jégou 等ICML 2023 · 被引用 114 次
- Idempotence and Perceptual Image CompressionTongda Xu, Ziran Zhu, Dailan He, Yanghao Li 等ICLR 2024 · 被引用 33 次
- One-Step Diffusion-Based Image Compression with Semantic DistillationNaifu Xue, Zhaoyang Jia, Jiahao Li, Bin Li 等NeurIPS 2025 · 被引用 28 次
它引用的顶会 Paper12
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li 等NeurIPS 2022 · 被引用 8,965 次
- Zero-Shot Text-to-Image GenerationAditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray 等ICML 2021 · 被引用 6,356 次
- High-Fidelity Generative Image CompressionFabian Mentzer, George Toderici, Michael Tschannen, Eirikur AgustssonNeurIPS 2020 · 被引用 675 次
- Generative Adversarial Networks for Extreme Learned Image CompressionEirikur Agustsson, Michael Tschannen, Fabian Mentzer, Radu Timofte 等ICCV 2019 · 被引用 648 次
相关 Paper
- Optimally Controllable Perceptual Lossy CompressionZeyu Yan, Fei Wen, Peilin LiuICML 2022 · 被引用 23 次
- Controllable Distortion-Perception Tradeoff Through Latent Diffusion for Neural Image CompressionChuqin Zhou, Guo Lu, Jiangchuan Li, Xiangyu Chen 等AAAI 2025 · 被引用 3 次
- On Perceptual Lossy Compression: The Cost of Perceptual Reconstruction and An Optimal Training FrameworkZeyu Yan, Fei Wen, Rendong Ying, Chao Ma 等ICML 2021 · 被引用 48 次
- Decouple Distortion from Perception: Region Adaptive Diffusion for Extreme-low Bitrate Perception Image CompressionJinchang Xu, Shaokang Wang, Jintao Chen, Zhe Li 等CVPR 2025
- Universal Rate-Distortion-Perception Representations for Lossy CompressionGeorge Zhang, Jingjing Qian, Jun Chen, Ashish KhistiNeurIPS 2021 · 被引用 108 次
