Proximal Mean-Field for Neural Network Quantization
Thalaiyasingam Ajanthan, Puneet K. Dokania, Richard Hartley, Philip H. S. Torr
Abstract
Compressing large Neural Networks (NN) by quantizing the parameters, while maintaining the performance is highly desirable due to reduced memory and time complexity. In this work, we cast NN quantization as a discrete labelling problem, and by examining relaxations, we design an efficient iterative optimization procedure that involves stochastic gradient descent followed by a projection. We prove that our simple projected gradient descent approach is, in fact, equivalent to a proximal version of the well-known mean-field method. These findings would allow the decades-old and theoretically grounded research on MRF optimization to be used to design better network quantization schemes. Our experiments on standard classification datasets (MNIST, CIFAR10/100, TinyImageNet) with convolutional and residual architectures show that our algorithm obtains fully-quantized networks with accuracies very close to the floating-point reference networks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 67b124f1-195d-4f51-9565-3a8c58aa2b08Cited by top-tier papers9
- Searching for Low-Bit Weights in Quantized Neural NetworksZhaohui Yang, Yunhe Wang, Kai Han, Chunjing Xu et al.NeurIPS 2020 · 103 citations
- Training Binary Neural Networks using the Bayesian Learning RuleXiangming Meng, Roman Bachmann, Mohammad Emtiyaz KhanICML 2020 · 47 citations
- UMEC: Unified model and embedding compression for efficient recommendation systemsJiayi Shen, Haotao Wang, Shupeng Gui, Jianchao Tan et al.ICLR 2021 · 23 citations
- Regularized Frank-Wolfe for Dense CRFs: Generalizing Mean Field and BeyondD. Khuê Lê-Huu, Karteek AlahariNeurIPS 2021 · 16 citations
- Fast and Efficient DNN Deployment via Deep Gaussian Transfer LearningQi Sun, Chen Bai, Tinghuan Chen, Hao Geng et al.ICCV 2021 · 7 citations
Related papers
- Data-Independent Neural Pruning via CoresetsBen Mussay, Margarita Osadchy, Vladimir Braverman, Samson Zhou et al.ICLR 2020 · 65 citations
- Optimal and Approximate Adaptive Stochastic QuantizationRan Ben-Basat, Yaniv Ben-Itzhak, Michael Mitzenmacher, Shay VargaftikNeurIPS 2024 · 12 citations
- Learning from Loss Landscape: Generalizable Mixed-Precision Quantization via Adaptive Sharpness-Aware Gradient AligningLianbo Ma, Jianlun Ma, Yuee Zhou, Guoyang Xie et al.ICML 2025
- And the Bit Goes Down: Revisiting the Quantization of Neural NetworksPierre Stock, Armand Joulin, Rémi Gribonval, Benjamin Graham et al.ICLR 2020 · 157 citations
- Demystifying and Generalizing BinaryConnectTim Dockhorn, Yaoliang Yu, Eyyüb Sari, Mahdi Zolnouri et al.NeurIPS 2021 · 14 citations
