Boosted Dynamic Neural Networks
Haichao Yu, Haoxiang Li, Gang Hua, Gao Huang, Humphrey Shi
Abstract
Early-exiting dynamic neural networks (EDNN), as one type of dynamic neural networks, has been widely studied recently. A typical EDNN has multiple prediction heads at different layers of the network backbone. During inference, the model will exit at either the last prediction head or an intermediate prediction head where the prediction confidence is higher than a predefined threshold. To optimize the model, these prediction heads together with the network backbone are trained on every batch of training data. This brings a train-test mismatch problem that all the prediction heads are optimized on all types of data in training phase while the deeper heads will only see difficult inputs in testing phase. Treating training and testing inputs differently at the two phases will cause the mismatch between training and testing data distributions. To mitigate this problem, we formulate an EDNN as an additive model inspired by gradient boosting, and propose multiple training techniques to optimize the model effectively. We name our method BoostNet. Our experiments show it achieves the state-of-the-art performance on CIFAR100 and ImageNet datasets in both anytime and budgeted-batch prediction modes. Our code is released at https://github.com/SHI-Labs/Boosted-Dynamic-Networks .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 18db48dd-6a0e-4494-a895-6022e58b6232Cited by top-tier papers8
- Jointly-Learned Exit and Inference for a Dynamic Neural NetworkFlorence Regol, Joud Chataoui, Mark CoatesICLR 2024 · 17 citations
- MA-Net: Rethinking Neural Unit in the Light of AstrocytesMengqiao Han, Liyuan Pan, Xiabi LiuAAAI 2024 · 5 citations
- Enhancing Adaptive Deep Networks for Image Classification via Uncertainty-aware Decision FusionXu Zhang, Zhipeng Xie, Haiyang Yu, Qitong Wang et al.ACM MM 2024 · 4 citations
- Is the acquisition worth the cost? Surrogate losses for Consistent Two-stage ClassifiersFlorence Regol, Joseph Cotnareanu, Theodore Glavas, Mark CoatesNeurIPS 2025 · 3 citations
- DarkDistill: Difficulty-Aligned Federated Early-Exit Network Training on Heterogeneous DevicesLehao Qu, Shuyuan Li, Zimu Zhou, Boyi Liu et al.KDD 2025
Builds on8
- Post-Training Quantization for Vision TransformerZhenhua Liu, Yunhe Wang, Kai Han, Wei Zhang et al.NeurIPS 2021 · 528 citations
- Glance and Focus: a Dynamic Approach to Reducing Spatial Redundancy in Image ClassificationYulin Wang, Kangchen Lv, Rui Huang, Shiji Song et al.NeurIPS 2020 · 179 citations
- Improved Techniques for Training Adaptive Deep NetworksHao Li, Hong Zhang, Xiaojuan Qi, Ruigang Yang et al.ICCV 2019 · 152 citations
- Any-Precision Deep Neural NetworksHaichao Yu, Haoxiang Li, Humphrey Shi, Thomas S. Huang et al.AAAI 2021 · 79 citations
- Collaboration of Experts: Achieving 80% Top-1 Accuracy on ImageNet with 100M FLOPsYikang Zhang, Zhuo Chen, Zhao ZhongICML 2022 · 11 citations
Related papers
- CEED: Collaborative Early Exit Neural Network Inference at the EdgeYichong Chen, Zifeng Niu, Manuel Roveri, Giuliano CasaleINFOCOM 2025 · 7 citations
- Distillation-Based Training for Multi-Exit ArchitecturesMary Phuong, Christoph LampertICCV 2019 · 205 citations
- Towards Anytime Classification in Early-Exit Architectures by Enforcing Conditional MonotonicityMetod Jazbec, James Urquhart Allingham, Dan Zhang, Eric T. NalisnickNeurIPS 2023 · 21 citations
- Harmonized Dense Knowledge Distillation Training for Multi-Exit ArchitecturesXinglu Wang, Yingming LiAAAI 2021 · 26 citations
- EENet: Energy Efficient Neural Networks with Run-time Power ManagementXiangjie Li, Yingtao Shen, An Zou, Yehan MaDAC 2023 · 6 citations
