Correcting Quantization-Induced Gradient Mismatch in Neural Image Compression
Changhao Peng, Yuqi Ye, Wei Gao
摘要
In recent years, neural image compression methods have achieved impressive performance in image compression tasks, most of which are based on variational auto-encoder with hyper-prior and autoregressive Gaussian entropy model. We first demonstrate that the way these end-to-end approaches handle quantization during training leads to a mismatch between the gradients direction of entropy model parameters (i.e., mean and standard deviation) and the direction they should be optimized towards during inference, making neural network difficult to learn accurate estimates of entropy model parameters. To address this issue, we then propose a two-step improvement: in the first step, use straight-through estimator to align the forward propagation during training with inference, thereby correcting the gradients of standard deviation parameters; in the second step, utilize gradients transfer that we propose and MSE-guided gradients to manually compensate for the gradients of mean parameters lost due to straight-through estimator. Finally, we also propose to freeze the auto-encoder and hyper auto-encoder in pre-trained models provided by existing works, and fine-tune only the modules that predict the entropy model parameters, enabling efficient validation of proposed improvements. Experimental results show that our improvements bring appreciable performance gains to state-of-the-art neural image compression models in recent years. Meanwhile, our improvements require no modification to the structure of pre-trained models and only lightweight fine-tuning, which shows strong plug-and-play capability and practical utility.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper16
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Transformer-based Transform CodingYinhao Zhu, Yang Yang, Taco CohenICLR 2022 · 被引用 218 次
- Enhanced Invertible Encoding for Learned Image CompressionYueqi Xie, Ka Leong Cheng, Qifeng ChenACM MM 2021 · 被引用 195 次
- MLIC: Multi-Reference Entropy Model for Learned Image CompressionWei Jiang, Jiayu Yang, Yongqi Zhai, Peirong Ning 等ACM MM 2023 · 被引用 117 次
相关 Paper
- Soft then Hard: Rethinking the Quantization in Neural Image CompressionZongyu Guo, Zhizheng Zhang, Runsen Feng, Zhibo ChenICML 2021 · 被引用 94 次
- Asymmetric Gained Deep Image Compression With Continuous Rate AdaptationZe Cui, Jing Wang, Shangyin Gao, Tiansheng Guo 等CVPR 2021
- Learning Accurate Entropy Model with Global Reference for Image CompressionYichen Qian, Zhiyu Tan, Xiuyu Sun, Ming Lin 等ICLR 2021 · 被引用 93 次
- Robustly overfitting latents for flexible neural image compressionYura Perugachi-Diaz, Arwin Gansekoele, Sandjai BhulaiNeurIPS 2024 · 被引用 5 次
- ICMH-Net: Neural Image Compression Towards both Machine Vision and Human VisionLei Liu, Zhihao Hu, Zhenghao Chen, Dong XuACM MM 2023 · 被引用 21 次
