Anytime Inference with Distilled Hierarchical Neural Ensembles
Adria Ruiz, Jakob Verbeek
Abstract
Inference in deep neural networks can be computationally expensive, and networks capable of anytime inference are important in scenarios where the amount of compute or quantity of input data varies over time. In such networks the inference process can interrupted to provide a result faster, or continued to obtain a more accurate result. We propose Hierarchical Neural Ensembles (HNE), a novel framework to embed an ensemble of multiple networks in a hierarchical tree structure, sharing intermediate layers. In HNE we control the complexity of inference on-the-fly by evaluating more or less models in the ensemble. Our second contribution is a novel hierarchical distillation method to boost the prediction accuracy of small ensembles. This approach leverages the nested structure of our ensembles, to optimally allocate accuracy and diversity across the individual models. Our experiments show that, compared to previous anytime inference models, HNE provides state-of-the-art accuracy-computate trade-offs on the CIFAR-10/100 and ImageNet datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext aa79e69e-144c-4c77-a0bb-ede0f5a0e40fCited by top-tier papers6
- Self-Contrastive Learning: Single-Viewed Supervised Contrastive Framework Using Sub-networkSangmin Bae, Sungnyun Kim, Jongwoo Ko, Gihun Lee et al.AAAI 2023 · 14 citations
- ProGMLP: A Progressive Framework for GNN-to-MLP Knowledge Distillation with Efficient Trade-offsWeigang Lu, Ziyu Guan, Wei Zhao, Yaming Yang et al.AAAI 2026 · 1 citation
- TIPS: Topologically Important Path Sampling for Anytime Neural NetworksGuihong Li, Kartikeya Bhardwaj, Yuedong Yang, Radu MarculescuICML 2023
- Neural Parameter Allocation SearchBryan A. Plummer, Nikoli Dryden, Julius Frost, Torsten Hoefler et al.ICLR 2022
- Progressive Ensemble Distillation: Building Ensembles for Efficient InferenceDon Kurian Dennis, Abhishek Shetty, Anish Prasad Sevekari, Kazuhito Koishida et al.NeurIPS 2023
Builds on8
- Searching for MobileNetV3Andrew Howard, Ruoming Pang, Hartwig Adam, Quoc V. Le et al.ICCV 2019 · 9,163 citations
- Once-for-All: Train One Network and Specialize it for Efficient DeploymentHan Cai, Chuang Gan, Tianzhe Wang, Zhekai Zhang et al.ICLR 2020 · 1,522 citations
- Ensemble Distribution DistillationAndrey Malinin, Bruno Mlodozeniec, Mark J. F. GalesICLR 2020 · 273 citations
- Depth-Adaptive TransformerMaha Elbayad, Jiatao Gu, Edouard Grave, Michael AuliICLR 2020 · 264 citations
- Improved Techniques for Training Adaptive Deep NetworksHao Li, Hong Zhang, Xiaojuan Qi, Ruigang Yang et al.ICCV 2019 · 152 citations
Related papers
- Adaptive Hierarchy-Branch Fusion for Online Knowledge DistillationLinrui Gong, Shaohui Lin, Baochang Zhang, Yunhang Shen et al.AAAI 2023 · 16 citations
- AnyDA: Anytime Domain AdaptationOmprakash Chakraborty, Aadarsh Sahoo, Rameswar Panda, Abir DasICLR 2023
- Improving Ensemble Distillation With Weight Averaging and Diversifying PerturbationGiung Nam, Hyungi Lee, Byeongho Heo, Juho LeeICML 2022 · 10 citations
- Distillation-Based Training for Multi-Exit ArchitecturesMary Phuong, Christoph LampertICCV 2019 · 205 citations
- Adaptive Depth Networks with Skippable Sub-PathsWoochul Kang, Hyungseop LeeNeurIPS 2024 · 5 citations
