Resilient Binary Neural Network
Sheng Xu, Yanjing Li, Teli Ma, Mingbao Lin, Hao Dong, Baochang Zhang, Peng Gao, Jinhu Lu
Abstract
Binary neural networks (BNNs) have received ever-increasing popularity for their great capability of reducing storage burden as well as quickening inference time. However, there is a severe performance drop compared with real-valued networks, due to its intrinsic frequent weight oscillation during training. In this paper, we introduce a Resilient Binary Neural Network (ReBNN) to mitigate the frequent oscillation for better BNNs' training. We identify that the weight oscillation mainly stems from the non-parametric scaling factor. To address this issue, we propose to parameterize the scaling factor and introduce a weighted reconstruction loss to build an adaptive training objective. For the first time, we show that the weight oscillation is controlled by the balanced parameter attached to the reconstruction loss, which provides a theoretical foundation to parameterize it in back propagation. Based on this, we learn our ReBNN by calculating the balanced parameter based on its maximum magnitude, which can effectively mitigate the weight oscillation with a resilient training process. Extensive experiments are conducted upon various network models, such as ResNet and Faster-RCNN for computer vision, as well as BERT for natural language processing. The results demonstrate the overwhelming performance of our ReBNN over prior arts. For example, our ReBNN achieves 66.9% Top-1 accuracy with ResNet-18 backbone on the ImageNet dataset, surpassing existing state-of-the-arts by a significant margin. Our code is open-sourced at https://github.com/SteveTsui/ReBNN .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f30edb2f-a7a5-4f3d-8e2b-577983b62853Cited by top-tier papers7
- DenseShift : Towards Accurate and Efficient Low-Bit Power-of-Two QuantizationXinlin Li, Bang Liu, Rui Heng Yang, Vanessa Courville et al.ICCV 2023 · 12 citations
- BiPFT: Binary Pre-trained Foundation Transformer with Low-Rank Estimation of Binarization Residual PolynomialsXingrun Xing, Li Du, Xinyuan Wang, Xianlin Zeng et al.AAAI 2024 · 5 citations
- Training Binary Neural Networks via Gaussian Variational Inference and Low-Rank Semidefinite ProgrammingLorenzo Orecchia, Jiawei Hu, Xue He, Wang Mark et al.NeurIPS 2024 · 4 citations
- Learning 1-Bit Tiny Object Detector with Discriminative Feature RefinementSheng Xu, Mingze Wang, Yanjing Li, Mingbao Lin et al.ICML 2024 · 4 citations
- BVT-IMA: Binary Vision Transformer with Information-Modified AttentionZhenyu Wang, Hao Luo, Xuemei Xie, Fan Wang et al.AAAI 2024 · 4 citations
Builds on9
- Q-BERT: Hessian Based Ultra Low Precision Quantization of BERTSheng Shen, Zhen Dong, Jiayu Ye, Linjian Ma et al.AAAI 2020 · 656 citations
- Training binary neural networks with real-to-binary convolutionsBrais Martínez, Jing Yang, Adrian Bulat, Georgios TzimiropoulosICLR 2020 · 251 citations
- Q-ViT: Accurate and Fully Quantized Low-bit Vision TransformerYanjing Li, Sheng Xu, Baochang Zhang, Xianbin Cao et al.NeurIPS 2022 · 185 citations
- Rotated Binary Neural NetworkMingbao Lin, Rongrong Ji, Zihan Xu, Baochang Zhang et al.NeurIPS 2020 · 161 citations
- TernaryBERT: Distillation-aware Ultra-low Bit BERTWei Zhang, Lu Hou, Yichun Yin, Lifeng Shang et al.EMNLP 2020 · 147 citations
Related papers
- How Do Adam and Training Strategies Help BNNs OptimizationZechun Liu, Zhiqiang Shen, Shichao Li, Koen Helwegen et al.ICML 2021 · 100 citations
- SA-BNN: State-Aware Binary Neural NetworkChunlei Liu, Peng Chen, Bohan Zhuang, Chunhua Shen et al.AAAI 2021 · 23 citations
- Understanding weight-magnitude hyperparameters in training binary networksJoris Quist, Yunqiang Li, Jan van GemertICLR 2023
- ReverB-SNN: Reversing Bit of the Weight and Activation for Spiking Neural NetworksYufei Guo, Yuhan Zhang, Jie Zhou, Xiaode Liu et al.ICML 2025
- Reducing the Computational Cost of Deep Generative Models with Binary Neural NetworksThomas Bird, Friso H. Kingma, David BarberICLR 2021 · 3 citations
