Scalable Model Compression by Entropy Penalized Reparameterization
Deniz Oktay, Johannes Ballé, Saurabh Singh, Abhinav Shrivastava
Abstract
We describe a simple and general neural network weight compression approach, in which the network parameters (weights and biases) are represented in a "latent" space, amounting to a reparameterization. This space is equipped with a learned probability model, which is used to impose an entropy penalty on the parameter representation during training, and to compress the representation using a simple arithmetic coder after training. Classification accuracy and model compressibility is maximized jointly, with the bitrate--accuracy trade-off specified by a hyperparameter. We evaluate the method on the MNIST, CIFAR-10 and ImageNet classification benchmarks using six distinct model architectures. Our results show that state-of-the-art model compression can be achieved in a scalable and general way without requiring complex procedures such as multi-stage training.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 5178d430-722e-4e26-b642-3e3c90b84097Cited by top-tier papers11
- The Fundamental Price of Secure Aggregation in Differentially Private Federated LearningWei-Ning Chen, Christopher A. Choquette-Choo, Peter Kairouz, Ananda Theertha SureshICML 2022 · 82 citations
- SHACIRA: Scalable HAsh-grid Compression for Implicit Neural RepresentationsSharath Girish, Abhinav Shrivastava, Kamal GuptaICCV 2023 · 36 citations
- Boosting Neural Representations for Videos with a Conditional DecoderXinjie Zhang, Ren Yang, Dailan He, Xingtong Ge et al.CVPR 2024 · 20 citations
- RECOMBINER: Robust and Enhanced Compression with Bayesian Implicit Neural RepresentationsJiajun He, Gergely Flamich, Zongyu Guo, José Miguel Hernández-LobatoICLR 2024 · 12 citations
- LilNetX: Lightweight Networks with EXtreme Model Compression and Structured SparsificationSharath Girish, Kamal Gupta, Saurabh Singh, Abhinav ShrivastavaICLR 2023 · 7 citations
Related papers
- Modality-Agnostic Variational Compression of Implicit Neural RepresentationsJonathan Richard Schwarz, Jihoon Tack, Yee Whye Teh, Jaeho Lee et al.ICML 2023 · 28 citations
- Compression with Bayesian Implicit Neural RepresentationsZongyu Guo, Gergely Flamich, Jiajun He, Zhibo Chen et al.NeurIPS 2023 · 38 citations
- Dynamic Model Pruning with FeedbackTao Lin, Sebastian U. Stich, Luis Barba, Daniil Dmitriev et al.ICLR 2020 · 229 citations
- Compressing Images by Encoding Their Latent Representations with Relative Entropy CodingGergely Flamich, Marton Havasi, José Miguel Hernández-LobatoNeurIPS 2020 · 78 citations
- Lossless Compression with Probabilistic CircuitsAnji Liu, Stephan Mandt, Guy Van den BroeckICLR 2022 · 29 citations
