Boosted Dynamic Neural Networks
Haichao Yu, Haoxiang Li, Gang Hua, Gao Huang, Humphrey Shi
摘要
Early-exiting dynamic neural networks (EDNN), as one type of dynamic neural networks, has been widely studied recently. A typical EDNN has multiple prediction heads at different layers of the network backbone. During inference, the model will exit at either the last prediction head or an intermediate prediction head where the prediction confidence is higher than a predefined threshold. To optimize the model, these prediction heads together with the network backbone are trained on every batch of training data. This brings a train-test mismatch problem that all the prediction heads are optimized on all types of data in training phase while the deeper heads will only see difficult inputs in testing phase. Treating training and testing inputs differently at the two phases will cause the mismatch between training and testing data distributions. To mitigate this problem, we formulate an EDNN as an additive model inspired by gradient boosting, and propose multiple training techniques to optimize the model effectively. We name our method BoostNet. Our experiments show it achieves the state-of-the-art performance on CIFAR100 and ImageNet datasets in both anytime and budgeted-batch prediction modes. Our code is released at https://github.com/SHI-Labs/Boosted-Dynamic-Networks .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Jointly-Learned Exit and Inference for a Dynamic Neural NetworkFlorence Regol, Joud Chataoui, Mark CoatesICLR 2024 · 被引用 17 次
- MA-Net: Rethinking Neural Unit in the Light of AstrocytesMengqiao Han, Liyuan Pan, Xiabi LiuAAAI 2024 · 被引用 5 次
- Enhancing Adaptive Deep Networks for Image Classification via Uncertainty-aware Decision FusionXu Zhang, Zhipeng Xie, Haiyang Yu, Qitong Wang 等ACM MM 2024 · 被引用 4 次
- Is the acquisition worth the cost? Surrogate losses for Consistent Two-stage ClassifiersFlorence Regol, Joseph Cotnareanu, Theodore Glavas, Mark CoatesNeurIPS 2025 · 被引用 3 次
- DarkDistill: Difficulty-Aligned Federated Early-Exit Network Training on Heterogeneous DevicesLehao Qu, Shuyuan Li, Zimu Zhou, Boyi Liu 等KDD 2025
它引用的顶会 Paper8
- Post-Training Quantization for Vision TransformerZhenhua Liu, Yunhe Wang, Kai Han, Wei Zhang 等NeurIPS 2021 · 被引用 528 次
- Glance and Focus: a Dynamic Approach to Reducing Spatial Redundancy in Image ClassificationYulin Wang, Kangchen Lv, Rui Huang, Shiji Song 等NeurIPS 2020 · 被引用 179 次
- Improved Techniques for Training Adaptive Deep NetworksHao Li, Hong Zhang, Xiaojuan Qi, Ruigang Yang 等ICCV 2019 · 被引用 152 次
- Any-Precision Deep Neural NetworksHaichao Yu, Haoxiang Li, Humphrey Shi, Thomas S. Huang 等AAAI 2021 · 被引用 79 次
- Collaboration of Experts: Achieving 80% Top-1 Accuracy on ImageNet with 100M FLOPsYikang Zhang, Zhuo Chen, Zhao ZhongICML 2022 · 被引用 11 次
相关 Paper
- CEED: Collaborative Early Exit Neural Network Inference at the EdgeYichong Chen, Zifeng Niu, Manuel Roveri, Giuliano CasaleINFOCOM 2025 · 被引用 7 次
- Distillation-Based Training for Multi-Exit ArchitecturesMary Phuong, Christoph LampertICCV 2019 · 被引用 205 次
- Towards Anytime Classification in Early-Exit Architectures by Enforcing Conditional MonotonicityMetod Jazbec, James Urquhart Allingham, Dan Zhang, Eric T. NalisnickNeurIPS 2023 · 被引用 21 次
- Harmonized Dense Knowledge Distillation Training for Multi-Exit ArchitecturesXinglu Wang, Yingming LiAAAI 2021 · 被引用 26 次
- EENet: Energy Efficient Neural Networks with Run-time Power ManagementXiangjie Li, Yingtao Shen, An Zou, Yehan MaDAC 2023 · 被引用 6 次
