Fast-BCNN: Massive Neuron Skipping in Bayesian Convolutional Neural Networks
Qiyu Wan, Xin Fu
Abstract
Bayesian Convolutional Neural Networks (BCNNs) have emerged as a robust form of Convolutional Neural Networks (CNNs) with the capability of uncertainty estimation. A BCNN model is implemented by adding a dropout layer after each convolutional layer in the original CNN. By executing the stochastic inferences many times, BCNNs are able to provide an output distribution that reflects the uncertainty of the final prediction. Repeated inferences in this process lead to much longer execution time, which makes it challenging to apply Bayesian technique to CNNs in real-world applications. In this study, we propose Fast-BCNN, an FPGA-based hardware accelerator design that intelligently skips the redundant computations for two types of neurons during repeated BCNN inferences. Firstly, within a sample inference, we aim to skip the dropped neurons that predetermined by dropout masks. Secondly, by leveraging the information from the first inference and dropout masks, we predict the zero neurons and skip all their corresponding computations during the following sample inferences. Particularly, an optimization algorithm is employed to guarantee the accuracy of zero neuron prediction while achieving the maximal computation reduction. To support our neuron skipping strategy at hardware level, we explore an efficient parallelism for CNN convolution to gracefully skip the corresponding computations for both types of neurons, we then propose a novel PE architecture that accommodates the parallel operation of convolution and prediction with negligible overhead. Experimental results demonstrate that our Fast-BCNN achieves 2.1 8.2× speedup and 44% 84% energy reduction over the baseline CNN accelerator.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Cited by top-tier papers3
- Shift-BNN: Highly-Efficient Probabilistic Bayesian Neural Network Training via Memory-Friendly Pattern RetrievingQiyu Wan, Haojun Xia, Xingyao Zhang, Lening Wang et al.MICRO 2021 · 9 citations
- Enabling fast uncertainty estimation: accelerating bayesian transformers via algorithmic and hardware optimizationsHongxiang Fan, Martin Ferianc, Wayne LukDAC 2022 · 7 citations
- When Monte-Carlo Dropout Meets Multi-Exit: Optimizing Bayesian Neural Networks on FPGAHongxiang Fan, Mark Chen, Liam Castelli, Zhiqiang Que et al.DAC 2023 · 5 citations
Related papers
- High-Performance FPGA-based Accelerator for Bayesian Neural NetworksHongxiang Fan, Martin Ferianc, Miguel Rodrigues, Hongyu Zhou et al.DAC 2021 · 31 citations
- Hardware-Aware Neural Dropout Search for Reliable Uncertainty Prediction on FPGAZehuan Zhang, Hongxiang Fan, Hao Mark Chen, Lukasz Dudziak et al.DAC 2024 · 1 citation
- Bayesian Nested Neural Networks for Uncertainty Calibration and Adaptive CompressionYufei Cui, Ziquan Liu, Qiao Li, Antoni B. Chan et al.CVPR 2021
- Accelerate CNN via Recursive Bayesian PruningYuefu Zhou, Ya Zhang, Yanfeng Wang, Qi TianICCV 2019 · 64 citations
- Towards Uncertainty-aware Robotic Perception via Mixed-signal BNN Engine Leveraging Probabilistic Quantum TunnelingLikai Pei, Yu Zhou, Xingtian Wang, Xueji Zhao et al.DAC 2025
