Prune and Tune Ensembles: Low-Cost Ensemble Learning with Sparse Independent Subnetworks
Tim Whitaker, Darrell Whitley
Abstract
Ensemble Learning is an effective method for improving generalization in machine learning. However, as state-of-the-art neural networks grow larger, the computational cost associated with training several independent networks becomes expensive. We introduce a fast, low-cost method for creating diverse ensembles of neural networks without needing to train multiple models from scratch. We do this by first training a single parent network. We then create child networks by cloning the parent and dramatically pruning the parameters of each child to create an ensemble of members with unique and diverse topologies. We then briefly train each child network for a small number of epochs, which now converge significantly faster when compared to training from scratch. We explore various ways to maximize diversity in the child networks, including the use of anti-random pruning and one-cycle tuning. This diversity enables "Prune and Tune" ensembles to achieve results that are competitive with traditional ensembles at a fraction of the training cost. We benchmark our approach against state of the art low-cost ensemble methods and display marked improvement in both accuracy and uncertainty estimation on CIFAR-10 and CIFAR-100.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3ed62f07-4e95-43bf-93bc-146cc52e82eeCited by top-tier papers4
- Sparse Model Soups: A Recipe for Improved Pruning via Model AveragingMax Zimmer, Christoph Spiegel, Sebastian PokuttaICLR 2024 · 22 citations
- Mastering Stock Markets with Efficient Mixture of Diversified Trading ExpertsShuo Sun, Xinrun Wang, Wanqi Xue, Xiaoxuan Lou et al.KDD 2023 · 13 citations
- Ensembling Pruned Attention Heads For Uncertainty-Aware Efficient TransformersFiras Gabetni, Giuseppe Curci, Andrea Pilzer, Subhankar Roy et al.ICLR 2026 · 5 citations
- Sharpness-diversity tradeoff: improving flat ensembles with SharpBalanceHaiquan Lu, Xiaotian Liu, Yefan Zhou, Qunli Li et al.NeurIPS 2024 · 4 citations
Builds on5
- BatchEnsemble: an Alternative Approach to Efficient Ensemble and Lifelong LearningYeming Wen, Dustin Tran, Jimmy BaICLR 2020 · 569 citations
- Hyperparameter Ensembles for Robustness and Uncertainty QuantificationFlorian Wenzel, Jasper Snoek, Dustin Tran, Rodolphe JenattonNeurIPS 2020 · 263 citations
- Training independent subnetworks for robust predictionMarton Havasi, Rodolphe Jenatton, Stanislav Fort, Jeremiah Zhe Liu et al.ICLR 2021 · 235 citations
- Network Pruning That Matters: A Case Study on Retraining VariantsDuong H. Le, Binh-Son HuaICLR 2021 · 45 citations
- Neural networks with late-phase weightsJohannes von Oswald, Seijin Kobayashi, João Sacramento, Alexander Meulemans et al.ICLR 2021 · 38 citations
Related papers
- Deep Ensembling with No Overhead for either Training or Testing: The All-Round Blessings of Dynamic SparsityShiwei Liu, Tianlong Chen, Zahra Atashgahi, Xiaohan Chen et al.ICLR 2022 · 62 citations
- Ensemble Pruning for Out-of-distribution GeneralizationFengchun Qiao, Xi PengICML 2024 · 3 citations
- Boost Neural Networks by CheckpointsFeng Wang, Guoyizhe Wei, Qiao Liu, Jinxiang Ou et al.NeurIPS 2021 · 13 citations
- Ex Uno Pluria: Insights on Ensembling in Low Precision Number SystemsGiung Nam, Juho LeeNeurIPS 2024 · 2 citations
- Masksembles for Uncertainty EstimationNikita Durasov, Timur M. Bagautdinov, Pierre Baqué, Pascal FuaCVPR 2021
