Reducing the Computational Cost of Deep Generative Models with Binary Neural Networks
Thomas Bird, Friso H. Kingma, David Barber
Abstract
Deep generative models provide a powerful set of tools to understand real-world data. But as these models improve, they increase in size and complexity, so their computational cost in memory and execution time grows. Using binary weights in neural networks is one method which has shown promise in reducing this cost. However, whether binary neural networks can be used in generative models is an open problem. In this work we show, for the first time, that we can successfully train generative models which utilize binary neural networks. This reduces the computational cost of the models massively. We develop a new class of binary weight normalization, and provide insights for architecture designs of these binarized generative models. We demonstrate that two state-of-the-art deep generative models, the ResNet VAE and Flow++ models, can be binarized effectively using these techniques. We train binary models that achieve loss values close to those of the regular models but are 90%-94% smaller in size, and also allow significant speed-ups in execution time.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6d215ed2-9d3b-4cdf-be12-8a6abf20ca6bCited by top-tier papers4
- On the Out-of-distribution Generalization of Probabilistic Image ModellingMingtian Zhang, Andi Zhang, Steven McDonaghNeurIPS 2021 · 51 citations
- Partition and Code: learning how to compress graphsGiorgos Bouritsas, Andreas Loukas, Nikolaos Karalias, Michael M. BronsteinNeurIPS 2021 · 23 citations
- BiDM: Pushing the Limit of Quantization for Diffusion ModelsXingyu Zheng, Xianglong Liu, Yichen Bian, Xudong Ma et al.NeurIPS 2024 · 12 citations
- Fast Lossless Neural Compression with Integer-Only Discrete FlowsSiyu Wang, Jianfei Chen, Chongxuan Li, Jun Zhu et al.ICML 2022 · 8 citations
Builds on1
Related papers
- DeepWeightFlow: Re-Basined Flow Matching for Generating Neural Network WeightsSaumya Gupta, Scott Biggs, Moritz Laber, Zohair Shafi et al.ICLR 2026 · 5 citations
- Resilient Binary Neural NetworkSheng Xu, Yanjing Li, Teli Ma, Mingbao Lin et al.AAAI 2023 · 1 citation
- Sub-bit Neural Networks: Learning to Compress and Accelerate Binary Neural NetworksYikai Wang, Yi Yang, Fuchun Sun, Anbang YaoICCV 2021 · 18 citations
- Understanding weight-magnitude hyperparameters in training binary networksJoris Quist, Yunqiang Li, Jan van GemertICLR 2023
- Fast and Accurate Binary Neural Networks Based on Depth-Width ReshapingPing Xue, Yang Lu, Jingfei Chang, Xing Wei et al.AAAI 2023 · 3 citations
