On Perceptual Lossy Compression: The Cost of Perceptual Reconstruction and An Optimal Training Framework
Zeyu Yan, Fei Wen, Rendong Ying, Chao Ma, Peilin Liu
摘要
Lossy compression algorithms are typically designed to achieve the lowest possible distortion at a given bit rate. However, recent studies show that pursuing high perceptual quality would lead to increase of the lowest achievable distortion (e.g., MSE). This paper provides nontrivial results theoretically revealing that, 1) the cost of achieving perfect perception quality is exactly a doubling of the lowest achievable MSE distortion, 2) an optimal encoder for the"classic"rate-distortion problem is also optimal for the perceptual compression problem, 3) distortion loss is unnecessary for training a perceptual decoder. Further, we propose a novel training framework to achieve the lowest MSE distortion under perfect perception constraint at a given bit rate. This framework uses a GAN with discriminator conditioned on an MSE-optimized encoder, which is superior over the traditional framework using distortion plus adversarial loss. Experiments are provided to verify the theoretical finding and demonstrate the superiority of the proposed training framework.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper12
- Improving Statistical Fidelity for Neural Image Compression with Implicit Local Likelihood ModelsMatthew J. Muckley, Alaaeldin El-Nouby, Karen Ullrich, Hervé Jégou 等ICML 2023 · 被引用 114 次
- Universal Rate-Distortion-Perception Representations for Lossy CompressionGeorge Zhang, Jingjing Qian, Jun Chen, Ashish KhistiNeurIPS 2021 · 被引用 108 次
- Idempotence and Perceptual Image CompressionTongda Xu, Ziran Zhu, Dailan He, Yanghao Li 等ICLR 2024 · 被引用 33 次
- Non-Semantics Suppressed Mask Learning for Unsupervised Video Semantic CompressionYuan Tian, Guo Lu, Guangtao Zhai, Zhiyong GaoICCV 2023 · 被引用 29 次
- Optimally Controllable Perceptual Lossy CompressionZeyu Yan, Fei Wen, Peilin LiuICML 2022 · 被引用 23 次
它引用的顶会 Paper2
相关 Paper
- Multi-Realism Image Compression with a Conditional GeneratorEirikur Agustsson, David Minnen, George Toderici, Fabian MentzerCVPR 2023
- On the choice of Perception Loss Function for Learned Video CompressionSadaf Salehkalaibar, Buu Phan, Jun Chen, Wei Yu 等NeurIPS 2023 · 被引用 23 次
- Video Coding using Learned Latent GAN CompressionMustafa Shukor, Bharath Bhushan Damodaran, Xu Yao, Pierre HellierACM MM 2022 · 被引用 6 次
- Training-Free Rate-Distortion-Perception Traversal With DiffusionYuhan Wang, Suzhi Bi, Angela Yingjun ZhangICML 2026
- Towards image compression with perfect realism at ultra-low bitratesMarlène Careil, Matthew J. Muckley, Jakob Verbeek, Stéphane LathuilièreICLR 2024 · 被引用 122 次
