BinaryDuo: Reducing Gradient Mismatch in Binary Activation Network by Coupling Binary Activations
Hyungjun Kim, Kyungsu Kim, Jinseok Kim, Jae-Joon Kim
Abstract
Binary Neural Networks (BNNs) have been garnering interest thanks to their compute cost reduction and memory savings. However, BNNs suffer from performance degradation mainly due to the gradient mismatch caused by binarizing activations. Previous works tried to address the gradient mismatch problem by reducing the discrepancy between activation functions used at forward pass and its differentiable approximation used at backward pass, which is an indirect measure. In this work, we use the gradient of smoothed loss function to better estimate the gradient mismatch in quantized neural network. Analysis using the gradient mismatch estimator indicates that using higher precision for activation is more effective than modifying the differentiable approximation of activation function. Based on the observation, we propose a new training scheme for binary activation networks called BinaryDuo in which two binary activations are coupled into a ternary activation during training. Experimental results show that BinaryDuo outperforms state-of-the-art BNNs on various benchmarks with the same amount of parameters and computing cost.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4b29e5ff-405e-4989-aeae-8d59aa3b5758Cited by top-tier papers5
- Spiking Neural Networks with Improved Inherent Recurrence Dynamics for Sequential LearningWachirawit Ponghiran, Kaushik RoyAAAI 2022 · 60 citations
- OMPQ: Orthogonal Mixed Precision QuantizationYuexiao Ma, Taisong Jin, Xiawu Zheng, Yan Wang et al.AAAI 2023 · 56 citations
- TRQ: Ternary Neural Networks With Residual QuantizationYue Li, Wenrui Ding, Chunlei Liu, Baochang Zhang et al.AAAI 2021 · 34 citations
- SA-BNN: State-Aware Binary Neural NetworkChunlei Liu, Peng Chen, Bohan Zhuang, Chunhua Shen et al.AAAI 2021 · 23 citations
- Fast and Accurate Binary Neural Networks Based on Depth-Width ReshapingPing Xue, Yang Lu, Jingfei Chang, Xing Wei et al.AAAI 2023 · 3 citations
Builds on1
Related papers
- BiPer: Binary Neural Networks Using a Periodic FunctionEdwin Vargas, Claudia V. Correa P., Carlos Hinojosa, Henry ArguelloCVPR 2024 · 10 citations
- SURGE: Surrogate Gradient Adaptation in Binary Neural NetworksHaoyu Huang, Boyu Liu, Linlin Yang, Yanjing Li et al.ICML 2026
- DIVISION: Memory Efficient Training via Dual Activation PrecisionGuanchu Wang, Zirui Liu, Zhimeng Jiang, Ninghao Liu et al.ICML 2023 · 4 citations
- Learning Frequency Domain Approximation for Binary Neural NetworksYixing Xu, Kai Han, Chang Xu, Yehui Tang et al.NeurIPS 2021 · 64 citations
- Rotated Binary Neural NetworkMingbao Lin, Rongrong Ji, Zihan Xu, Baochang Zhang et al.NeurIPS 2020 · 161 citations
