Ensemble Pruning for Out-of-distribution Generalization
Fengchun Qiao, Xi Peng
Abstract
Ensemble of deep neural networks has achieved great success in hedging against single-model failure under distribution shift. However, existing techniques suffer from producing redundant models, limiting predictive diversity and yielding compromised generalization performance. Existing ensemble pruning methods can only guarantee predictive diversity for in-distribution data, which may not transfer well to out-of-distribution (OoD) data. To address this gap, we propose a principled optimization framework for ensemble pruning under distribution shifts. Since the annotations of test data are not available, we explore relationships between prediction distributions of the models, encapsulated in a topology graph. By incorporating this topology into a combinatorial optimization framework, complementary models with high predictive diversity are selected with theoretical guarantees. Our approach is model-agnostic and can be applied on top of a broad spectrum of off-the-shelf ensembling methods for improved generalization performance. Experiments on common benchmarks demonstrate the superiority of our approach in both multi-and single-source OoD generalization. The source codes are publicly available at: https://github.com/joffery/TEP .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b22c668a-e59e-49ed-9b03-4bfaf813a329Cited by top-tier papers2
- Supporting Multimodal Intermediate Fusion with Informatic Constraint and Distribution CoherenceYi Li, Fei Song, Changwen Zheng, Jiangmeng LiICLR 2026
- Structure-informed Risk Minimization for Robust Ensemble LearningFengchun Qiao, Yanlin Chen, Xi PengICML 2025
Builds on20
- WILDS: A Benchmark of in-the-Wild Distribution ShiftsPang Wei Koh, Shiori Sagawa, Henrik Marklund, Sang Michael Xie et al.ICML 2021 · 1,773 citations
- Distributionally Robust Neural NetworksShiori Sagawa, Pang Wei Koh, Tatsunori B. Hashimoto, Percy LiangICLR 2020 · 1,578 citations
- Model soups: averaging weights of multiple fine-tuned models improves accuracy without increasing inference timeMitchell Wortsman, Gabriel Ilharco, Samir Yitzhak Gadre, Rebecca Roelofs et al.ICML 2022 · 1,464 citations
- In Search of Lost Domain GeneralizationIshaan Gulrajani, David Lopez-PazICLR 2021 · 1,416 citations
- Test-Time Training with Self-Supervision for Generalization under Distribution ShiftsYu Sun, Xiaolong Wang, Zhuang Liu, John Miller et al.ICML 2020 · 1,220 citations
Related papers
- DIBS: Diversity Inducing Information Bottleneck in Model EnsemblesSamarth Sinha, Homanga Bharadhwaj, Anirudh Goyal, Hugo Larochelle et al.AAAI 2021 · 43 citations
- Pruning Spurious Subgraphs for Graph Out-of-Distribution GeneralizationTianjun Yao, Haoxuan Li, Yongqiang Chen, Tongliang Liu et al.NeurIPS 2025
- Robustness via Cross-Domain EnsemblesTeresa Yeo, Oguzhan Fatih Kar, Amir ZamirICCV 2021 · 30 citations
- Protecting DNNs from Theft using an Ensemble of Diverse ModelsSanjay Kariyappa, Atul Prakash, Moinuddin K. QureshiICLR 2021 · 33 citations
- Ensembles of Locally Independent Prediction ModelsAndrew Slavin Ross, Weiwei Pan, Leo A. Celi, Finale Doshi-VelezAAAI 2020 · 33 citations
