Training binary neural networks with real-to-binary convolutions
Brais Martínez, Jing Yang, Adrian Bulat, Georgios Tzimiropoulos
摘要
This paper shows how to train binary networks to within a few percent points ( 3-5 %) of the full precision counterpart with a negligible increase in the computational cost. In particular, we first show how to build a strong baseline, which already achieves state-of-the-art accuracy, by combining recently proposed advances, and carefully tuning the optimization procedure. Secondly, we show that by attempting to minimize the discrepancy between the output of the binary and the corresponding real-valued convolution additional significant accuracy gains can be obtained. We materialize this idea in two complementary ways: (1) with a loss function, during training, by matching the spatial attention maps computed at the output of the binary and real-valued convolutions, and (2) in data-driven manner, by using the real-valued activations being available during inference prior to the binarization process for re-scaling the activations right after the binary convolution. Finally, we show that, when putting all of our improvements together, the resulting model reduces the gap to its real-valued counterpart to less than 3% and 5% top-1 error on CIFAR-100 and ImageNet, respectively, when using a ResNet-18 architecture.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper27
- BiBERT: Accurate Fully Binarized BERTHaotong Qin, Yifu Ding, Mingyuan Zhang, Qinghua Yan 等ICLR 2022 · 被引用 121 次
- How Do Adam and Training Strategies Help BNNs OptimizationZechun Liu, Zhiqiang Shen, Shichao Li, Koen Helwegen 等ICML 2021 · 被引用 100 次
- BiT: Robustly Binarized Multi-distilled TransformerZechun Liu, Barlas Oguz, Aasish Pappu, Lin Xiao 等NeurIPS 2022 · 被引用 93 次
- Is Label Smoothing Truly Incompatible with Knowledge Distillation: An Empirical StudyZhiqiang Shen, Zechun Liu, Dejia Xu, Zitian Chen 等ICLR 2021 · 被引用 83 次
- BiPointNet: Binary Neural Network for Point CloudsHaotong Qin, Zhongang Cai, Mingyuan Zhang, Yifu Ding 等ICLR 2021 · 被引用 54 次
它引用的顶会 Paper1
相关 Paper
- High-Capacity Expert Binary NetworksAdrian Bulat, Brais Martínez, Georgios TzimiropoulosICLR 2021 · 被引用 29 次
- Sparsity-Inducing Binarized Neural NetworksPeisong Wang, Xiangyu He, Gang Li, Tianli Zhao 等AAAI 2020 · 被引用 60 次
- Fast and Accurate Binary Neural Networks Based on Depth-Width ReshapingPing Xue, Yang Lu, Jingfei Chang, Xing Wei 等AAAI 2023 · 被引用 3 次
- Training Binary Neural Network without Batch Normalization for Image Super-ResolutionXinrui Jiang, Nannan Wang, Jingwei Xin, Keyu Li 等AAAI 2021 · 被引用 52 次
- SA-BNN: State-Aware Binary Neural NetworkChunlei Liu, Peng Chen, Bohan Zhuang, Chunhua Shen 等AAAI 2021 · 被引用 23 次
