Improving Lossless Compression Rates via Monte Carlo Bits-Back Coding
Yangjun Ruan, Karen Ullrich, Daniel Severo, James Townsend, Ashish Khisti, Arnaud Doucet, Alireza Makhzani, Chris J. Maddison
Abstract
Latent variable models have been successfully applied in lossless compression with the bits-back coding algorithm. However, bits-back suffers from an increase in the bitrate equal to the KL divergence between the approximate posterior and the true posterior. In this paper, we show how to remove this gap asymptotically by deriving bits-back coding algorithms from tighter variational bounds. The key idea is to exploit extended space representations of Monte Carlo estimators of the marginal likelihood. Naively applied, our schemes would require more initial bits than the standard bits-back coder, but we show how to drastically reduce this additional cost with couplings in the latent space. When parallel architectures can be exploited, our coders can achieve better rates than bits-back with little additional cost. We demonstrate improved lossless compression rates in a variety of settings, especially in out-of-distribution or sequential data compression.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9aa3ff16-844a-4fdb-a86a-6b86758ec17bCited by top-tier papers9
- Differentiable Annealed Importance Sampling and the Perils of Gradient NoiseGuodong Zhang, Kyle Hsu, Jianing Li, Chelsea Finn et al.NeurIPS 2021 · 46 citations
- Lossless Compression with Probabilistic CircuitsAnji Liu, Stephan Mandt, Guy Van den BroeckICLR 2022 · 29 citations
- Sparse Probabilistic Circuits via Pruning and GrowingMeihua Dang, Anji Liu, Guy Van den BroeckNeurIPS 2022 · 25 citations
- Partition and Code: learning how to compress graphsGiorgos Bouritsas, Andreas Loukas, Nikolaos Karalias, Michael M. BronsteinNeurIPS 2021 · 23 citations
- Generalization Gap in Amortized InferenceMingtian Zhang, Peter Hayes, David BarberNeurIPS 2022 · 14 citations
Builds on5
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- NVAE: A Deep Hierarchical Variational AutoencoderArash Vahdat, Jan KautzNeurIPS 2020 · 1,141 citations
- Compressing Images by Encoding Their Latent Representations with Relative Entropy CodingGergely Flamich, Marton Havasi, José Miguel Hernández-LobatoNeurIPS 2020 · 78 citations
- HiLLoC: lossless image compression with hierarchical latent variable modelsJames Townsend, Thomas Bird, Julius Kunze, David BarberICLR 2020 · 60 citations
- IDF++: Analyzing and Improving Integer Discrete Flows for Lossless CompressionRianne van den Berg, Alexey A. Gritsenko, Mostafa Dehghani, Casper Kaae Sønderby et al.ICLR 2021 · 38 citations
Related papers
- Improving Inference for Neural Image CompressionYibo Yang, Robert Bamler, Stephan MandtNeurIPS 2020 · 151 citations
- SUMO: Unbiased Estimation of Log Marginal Probability for Latent Variable ModelsYucen Luo, Alex Beatson, Mohammad Norouzi, Jun Zhu et al.ICLR 2020 · 29 citations
- Compressing Tabular Data via Latent Variable EstimationAndrea Montanari, Eric WeinerICML 2023
- Learned Lossless Image Compression Based on Bit Plane SlicingZhe Zhang, Huairui Wang, Zhenzhong Chen, Shan LiuCVPR 2024 · 4 citations
- Amortised Learning by Wake-SleepLi K. Wenliang, Theodore H. Moskovitz, Heishiro Kanagawa, Maneesh SahaniICML 2020 · 7 citations
