Good, Cheap, and Fast: Overfitted Image Compression with Wasserstein Distortion
Jona Ballé, Luca Versari, Emilien Dupont, Hyunjik Kim, Matthias Bauer
摘要
Inspired by the success of generative image models, recent work on learned image compression increasingly focuses on better probabilistic models of the natural image distribution, leading to excellent image quality. This, however, comes at the expense of a computational complexity that is several orders of magnitude higher than today's commercial codecs, and thus prohibitive for most practical applications. With this paper, we demonstrate that by focusing on modeling visual perception rather than the data distribution, we can achieve a very good trade-off between visual quality and bit rate similar to "generative" compression models such as HiFiC, while requiring less than 1% of the multiply-accumulate operations (MACs) for decompression. We do this by optimizing C3, an overfitted image codec, for Wasserstein Distortion (WD), and evaluating the image reconstructions with a human rater study, showing that WD clearly outperforms LPIPS as an optimization objective. The study also reveals that WD outperforms other perceptual metrics such as LPIPS, DISTS, and MS-SSIM as a predictor of human ratings, remarkably achieving over 94% Pearson correlation with Elo scores.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- MoRIC: A Modular Region-based Implicit Codec for Image CompressionGen Li, Haotian Wu, Deniz GündüzNeurIPS 2025 · 被引用 5 次
- DiT-IC: Aligned Diffusion Transformer for Efficient Image CompressionJunqi Shi, Ming Lu, Xingchen Li, Anle Ke 等CVPR 2026 · 被引用 4 次
- What Matters in Practical Learned Image CompressionKedar Tatwawadi, Parisa Rahimzadeh, Zhanghao Sun, Zhiqi Chen 等CVPR 2026 · 被引用 3 次
- DynaQuant: Dynamic Mixed-Precision Quantization for Learned Image CompressionYouneng Bao, Yulong Cheng, Yiping Liu, Yichen Yang 等AAAI 2026
它引用的顶会 Paper17
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- High-Fidelity Generative Image CompressionFabian Mentzer, George Toderici, Michael Tschannen, Eirikur AgustssonNeurIPS 2020 · 被引用 675 次
- Lossy Image Compression with Conditional Diffusion ModelsRuihan Yang, Stephan MandtNeurIPS 2023 · 被引用 268 次
- HiNeRV: Video Compression with Hierarchical Encoding-based Neural RepresentationHo Man Kwan, Ge Gao, Fan Zhang, Andrew Gower 等NeurIPS 2023 · 被引用 132 次
- MLIC: Multi-Reference Entropy Model for Learned Image CompressionWei Jiang, Jiayu Yang, Yongqi Zhai, Peirong Ning 等ACM MM 2023 · 被引用 117 次
相关 Paper
- Frequency-Aware Perceptual Optimization for Low-Complexity Implicit Image CompressionHaotian Wu, Gen Li, Di You, Pier Luigi Dragotti 等ICML 2026
- C3: High-Performance and Low-Complexity Neural Compression from a Single Image or VideoHyunjik Kim, Matthias Bauer, Lucas Theis, Jonathan Richard Schwarz 等CVPR 2024
- How To Exploit the Transferability of Learned Image Compression to Conventional CodecsJan P. Klopp, Keng-Chi Liu, Liang-Gee Chen, Shao-Yi ChienCVPR 2021
- Improving Statistical Fidelity for Neural Image Compression with Implicit Local Likelihood ModelsMatthew J. Muckley, Alaaeldin El-Nouby, Karen Ullrich, Hervé Jégou 等ICML 2023 · 被引用 114 次
- Generative Latent Coding for Ultra-Low Bitrate Image CompressionZhaoyang Jia, Jiahao Li, Bin Li, Houqiang Li 等CVPR 2024
