Handling Long-tailed Feature Distribution in AdderNets
Minjing Dong, Yunhe Wang, Xinghao Chen, Chang Xu
Abstract
Adder neural networks (ANNs) are designed for low energy cost which replace expensive multiplications in convolutional neural networks (CNNs) with cheaper additions to yield energy-efficient neural networks and hardware accelerations. Although ANNs achieve satisfactory efficiency, there exist gaps between ANNs and CNNs where the accuracy of ANNs can hardly be compared to CNNs without the assistance of other training tricks, such as knowledge distillation. The inherent discrepancy lies in the similarity measurement between filters and features, however how to alleviate this difference remains unexplored. To locate the potential problem of ANNs, we focus on the property difference due to similarity measurement. We demonstrate that unordered heavy tails in ANNs could be the key component which prevents ANNs from achieving superior classification performance since fatter tails tend to overlap in feature space. Through pre-defining Multivariate Skew Laplace distributions and embedding feature distributions into the loss function, ANN features can be fully controlled and designed for various properties. We further present a novel method for tackling existing heavy tails in ANNs with only a modification of classifier where ANN features are clustered with their tails wellformulated through proposed angle-based constraint on the distribution parameters to encourage high diversity of tails. Experiments conducted on several benchmarks and comparison with other distributions demonstrate the effectiveness of proposed approach for boosting the performance of ANNs.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 824fc8c6-e67e-4416-b45a-3f6f38ce10e4Cited by top-tier papers2
- An Empirical Study of Adder Neural Networks for Object DetectionXinghao Chen, Chang Xu, Minjing Dong, Chunjing Xu et al.NeurIPS 2021 · 22 citations
- Neural Architecture RetrievalXiaohuan Pei, Yanxi Li, Minjing Dong, Chang XuICLR 2024
Builds on6
- ShiftAddNet: A Hardware-Inspired Deep NetworkHaoran You, Xiaohan Chen, Yongan Zhang, Chaojian Li et al.NeurIPS 2020 · 99 citations
- Kernel Based Progressive Distillation for Adder Neural NetworksYixing Xu, Chang Xu, Xinghao Chen, Wei Zhang et al.NeurIPS 2020 · 48 citations
- CARS: Continuous Evolution for Efficient Neural Architecture SearchZhaohui Yang, Yunhe Wang, Xinghao Chen, Boxin Shi et al.CVPR 2020
- AdderNet: Do We Really Need Multiplications in Deep Learning?Hanting Chen, Yunhe Wang, Chunjing Xu, Boxin Shi et al.CVPR 2020
- Manifold Regularized Dynamic Network PruningYehui Tang, Yunhe Wang, Yixing Xu, Yiping Deng et al.CVPR 2021
Related papers
- Towards Stable and Robust AdderNetsMinjing Dong, Yunhe Wang, Xinghao Chen, Chang XuNeurIPS 2021 · 11 citations
- Redistribution of Weights and Activations for AdderNet QuantizationYing Nie, Kai Han, Haikang Diao, Chuanjian Liu et al.NeurIPS 2022 · 14 citations
- Exploring Salient Object Detection with Adder Neural NetworksBo-Wen Yin, Zheng LinAAAI 2025 · 4 citations
- AdderSR: Towards Energy Efficient Image Super-ResolutionDehua Song, Yunhe Wang, Hanting Chen, Chang Xu et al.CVPR 2021
- EnOF-SNN: Training Accurate Spiking Neural Networks via Enhancing the Output FeatureYufei Guo, Weihang Peng, Xiaode Liu, Yuanpei Chen et al.NeurIPS 2024 · 21 citations
