Overfitting for Fun and Profit: Instance-Adaptive Data Compression
Ties van Rozendaal, Iris A. M. Huijben, Taco Cohen
Abstract
Neural data compression has been shown to outperform classical methods in terms of rate-distortion (RD) performance, with results still improving rapidly. At a high level, neural compression is based on an autoencoder that tries to reconstruct the input instance from a (quantized) latent representation, coupled with a prior that is used to losslessly compress these latents. Due to limitations on model capacity and imperfect optimization and generalization, such models will suboptimally compress test data in general. However, one of the great strengths of learned compression is that if the test-time data distribution is known and relatively lowentropy (e.g. a camera watching a static scene, a dash cam in an autonomous car, etc.), the model can easily be finetuned or adapted to this distribution, leading to improved RD performance. In this paper we take this concept to the extreme, adapting the full model to a single video, and sending model updates (quantized and compressed using a parameter-space prior) along with the latent representation. Unlike previous work, we finetune not only the encoder/latents but the entire model, and -during finetuning -take into account both the effect of model quantization and the additional costs incurred by sending the model updates. We evaluate an image compression model on I-frames (sampled at 2 fps) from videos of the Xiph dataset, and demonstrate that full-model adaptation improves RD performance by ∼ 1 dB, with respect to encoder-only finetuning.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9e2f453c-859d-40ba-9316-17ca61154eeeCited by top-tier papers16
- Transformer-based Transform CodingYinhao Zhu, Yang Yang, Taco CohenICLR 2022 · 218 citations
- NVRC: Neural Video Representation CompressionHo Man Kwan, Ge Gao, Fan Zhang, Andrew Gower et al.NeurIPS 2024 · 44 citations
- Computationally-Efficient Neural Image Compression with Shallow DecodersYibo Yang, Stephan MandtICCV 2023 · 43 citations
- Bit Allocation using OptimizationTongda Xu, Han Gao, Chenjian Gao, Yuanyuan Wang et al.ICML 2023 · 23 citations
- Dec-Adapter: Exploring Efficient Decoder-Side Adapter for Bridging Screen Content and Natural Image CompressionSheng Shen, Huanjing Yue, Jingyu YangICCV 2023 · 23 citations
Builds on6
- Learned Video CompressionOren Rippel, Sanjay Nair, Carissa Lew, Steve Branson et al.ICCV 2019 · 258 citations
- Video Compression With Rate-Distortion AutoencodersAmirHossein Habibian, Ties van Rozendaal, Jakub M. Tomczak, Taco CohenICCV 2019 · 233 citations
- Neural Inter-Frame Compression for Video CodingAbdelaziz Djelouah, Joaquim Campos, Simone Schaub-Meyer, Christopher SchroersICCV 2019 · 207 citations
- Improving Inference for Neural Image CompressionYibo Yang, Robert Bamler, Stephan MandtNeurIPS 2020 · 151 citations
- Learned Video Compression via Joint Spatial-Temporal Correlation ExplorationHaojie Liu, Han Shen, Lichao Huang, Ming Lu et al.AAAI 2020 · 63 citations
Related papers
- Video Compression with Entropy-Constrained Neural RepresentationsCarlos Gomes, Roberto Azevedo, Christopher SchroersCVPR 2023
- NeRFCodec: Neural Feature Compression Meets Neural Radiance Fields for Memory-Efficient Scene RepresentationSicheng Li, Hao Li, Yiyi Liao, Lu YuCVPR 2024
- Online Learned Continual Compression with Adaptive Quantization ModulesLucas Caccia, Eugene Belilovsky, Massimo Caccia, Joelle PineauICML 2020 · 95 citations
- Rate-aware Compression for NeRF-based Volumetric VideoZhiyu Zhang, Guo Lu, Huanxiong Liang, Zhengxue Cheng et al.ACM MM 2024 · 3 citations
- Slimmable Compressive Autoencoders for Practical Neural Image CompressionFei Yang, Luis Herranz, Yongmei Cheng, Mikhail G. MozerovCVPR 2021
