Scalable Model Compression by Entropy Penalized Reparameterization
Deniz Oktay, Johannes Ballé, Saurabh Singh, Abhinav Shrivastava
摘要
We describe a simple and general neural network weight compression approach, in which the network parameters (weights and biases) are represented in a "latent" space, amounting to a reparameterization. This space is equipped with a learned probability model, which is used to impose an entropy penalty on the parameter representation during training, and to compress the representation using a simple arithmetic coder after training. Classification accuracy and model compressibility is maximized jointly, with the bitrate--accuracy trade-off specified by a hyperparameter. We evaluate the method on the MNIST, CIFAR-10 and ImageNet classification benchmarks using six distinct model architectures. Our results show that state-of-the-art model compression can be achieved in a scalable and general way without requiring complex procedures such as multi-stage training.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- The Fundamental Price of Secure Aggregation in Differentially Private Federated LearningWei-Ning Chen, Christopher A. Choquette-Choo, Peter Kairouz, Ananda Theertha SureshICML 2022 · 被引用 82 次
- SHACIRA: Scalable HAsh-grid Compression for Implicit Neural RepresentationsSharath Girish, Abhinav Shrivastava, Kamal GuptaICCV 2023 · 被引用 36 次
- Boosting Neural Representations for Videos with a Conditional DecoderXinjie Zhang, Ren Yang, Dailan He, Xingtong Ge 等CVPR 2024 · 被引用 20 次
- RECOMBINER: Robust and Enhanced Compression with Bayesian Implicit Neural RepresentationsJiajun He, Gergely Flamich, Zongyu Guo, José Miguel Hernández-LobatoICLR 2024 · 被引用 12 次
- LilNetX: Lightweight Networks with EXtreme Model Compression and Structured SparsificationSharath Girish, Kamal Gupta, Saurabh Singh, Abhinav ShrivastavaICLR 2023 · 被引用 7 次
相关 Paper
- Modality-Agnostic Variational Compression of Implicit Neural RepresentationsJonathan Richard Schwarz, Jihoon Tack, Yee Whye Teh, Jaeho Lee 等ICML 2023 · 被引用 28 次
- Compression with Bayesian Implicit Neural RepresentationsZongyu Guo, Gergely Flamich, Jiajun He, Zhibo Chen 等NeurIPS 2023 · 被引用 38 次
- Dynamic Model Pruning with FeedbackTao Lin, Sebastian U. Stich, Luis Barba, Daniil Dmitriev 等ICLR 2020 · 被引用 229 次
- Compressing Images by Encoding Their Latent Representations with Relative Entropy CodingGergely Flamich, Marton Havasi, José Miguel Hernández-LobatoNeurIPS 2020 · 被引用 78 次
- Lossless Compression with Probabilistic CircuitsAnji Liu, Stephan Mandt, Guy Van den BroeckICLR 2022 · 被引用 29 次
