Lossy Compression for Lossless Prediction
Yann Dubois, Benjamin Bloem-Reddy, Karen Ullrich, Chris J. Maddison
Abstract
Most data is automatically collected and only ever"seen"by algorithms. Yet, data compressors preserve perceptual fidelity rather than just the information needed by algorithms performing downstream tasks. In this paper, we characterize the bit-rate required to ensure high performance on all predictive tasks that are invariant under a set of transformations, such as data augmentations. Based on our theory, we design unsupervised objectives for training neural compressors. Using these objectives, we train a generic image compressor that achieves substantial rate savings (more than on ImageNet) compared to JPEG on 8 datasets, without decreasing downstream classification performance.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d8aa5245-d902-4021-9a94-ce31d4b155c5Cited by top-tier papers17
- SatMAE: Pre-training Transformers for Temporal and Multi-Spectral Satellite ImageryYezhen Cong, Samar Khanna, Chenlin Meng, Patrick Liu et al.NeurIPS 2022 · 707 citations
- Improving Statistical Fidelity for Neural Image Compression with Implicit Local Likelihood ModelsMatthew J. Muckley, Alaaeldin El-Nouby, Karen Ullrich, Hervé Jégou et al.ICML 2023 · 114 citations
- Optimal Representations for Covariate ShiftYangjun Ruan, Yann Dubois, Chris J. MaddisonICLR 2022 · 77 citations
- Information-theoretic Online Memory Selection for Continual LearningShengyang Sun, Daniele Calandriello, Huiyi Hu, Ang Li et al.ICLR 2022 · 61 citations
- Compressive Visual RepresentationsKuang-Huei Lee, Anurag Arnab, Sergio Guadarrama, John F. Canny et al.NeurIPS 2021 · 55 citations
Builds on18
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Supervised Contrastive LearningPrannay Khosla, Piotr Teterwak, Chen Wang, Aaron Sarna et al.NeurIPS 2020 · 7,049 citations
- Barlow Twins: Self-Supervised Learning via Redundancy ReductionJure Zbontar, Li Jing, Ishan Misra, Yann LeCun et al.ICML 2021 · 2,942 citations
- VICReg: Variance-Invariance-Covariance Regularization for Self-Supervised LearningAdrien Bardes, Jean Ponce, Yann LeCunICLR 2022 · 1,226 citations
Related papers
- Bridging Compressed Image Latents and Multimodal Large Language ModelsChia-Hao Kao, Cheng Chien, Yu-Jen Tseng, Yi-Hsin Chen et al.ICLR 2025
- On Perceptual Lossy Compression: The Cost of Perceptual Reconstruction and An Optimal Training FrameworkZeyu Yan, Fei Wen, Rendong Ying, Chao Ma et al.ICML 2021 · 48 citations
- Generalization Gap in Amortized InferenceMingtian Zhang, Peter Hayes, David BarberNeurIPS 2022 · 14 citations
- Learning Optimal Representations with the Decodable Information BottleneckYann Dubois, Douwe Kiela, David J. Schwab, Ramakrishna VedantamNeurIPS 2020 · 58 citations
- Universal Rate-Distortion-Perception Representations for Lossy CompressionGeorge Zhang, Jingjing Qian, Jun Chen, Ashish KhistiNeurIPS 2021 · 108 citations
