Structure-informed Risk Minimization for Robust Ensemble Learning
Fengchun Qiao, Yanlin Chen, Xi Peng
Abstract
Ensemble learning is a powerful approach for improving generalization under distribution shifts, but its effectiveness heavily depends on how individual models are combined. Existing methods often optimize ensemble weights based on validation data, which may not represent unseen test distributions, leading to suboptimal performance in out-of-distribution (OoD) settings. Inspired by Distributionally Robust Optimization (DRO), we propose Structure-informed Risk Minimization (SRM), a principled framework that learns robust ensemble weights without access to test data. Unlike standard DRO, which defines uncertainty sets based on divergence metrics alone, SRM incorporates structural information of training distributions, ensuring that the uncertainty set aligns with plausible real-world shifts. This approach mitigates the over-pessimism of traditional worst-case optimization while maintaining robustness. We introduce a computationally efficient optimization algorithm with theoretical guarantees and demonstrate that SRM achieves superior OoD generalization compared to existing ensemble combination strategies across diverse benchmarks. Code is available at: https: //github.com/deep-real/SRM .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e5252eb1-fb7b-47d7-9706-878dceeed3a1Cited by top-tier papers1
Ask how each one uses itBuilds on22
- Moment Matching for Multi-Source Domain AdaptationXingchao Peng, Qinxun Bai, Xide Xia, Zijun Huang et al.ICCV 2019 · 2,239 citations
- WILDS: A Benchmark of in-the-Wild Distribution ShiftsPang Wei Koh, Shiori Sagawa, Henrik Marklund, Sang Michael Xie et al.ICML 2021 · 1,773 citations
- Distributionally Robust Neural NetworksShiori Sagawa, Pang Wei Koh, Tatsunori B. Hashimoto, Percy LiangICLR 2020 · 1,578 citations
- Model soups: averaging weights of multiple fine-tuned models improves accuracy without increasing inference timeMitchell Wortsman, Gabriel Ilharco, Samir Yitzhak Gadre, Rebecca Roelofs et al.ICML 2022 · 1,464 citations
- In Search of Lost Domain GeneralizationIshaan Gulrajani, David Lopez-PazICLR 2021 · 1,416 citations
Related papers
- Multi-Expert Distributionally Robust Optimization for Out-of-Distribution GeneralizationJinyong Jeong, Hyungu Kahng, Seoung Bum KimNeurIPS 2025 · 6 citations
- Causal Structure-guided Distributionally Robust Optimization under Domain ShiftsSeonggyeom Kim, Eunjung Choi, Dong-Kyu ChaeKDD 2026
- Sufficient Invariant Learning for Distribution ShiftTaero Kim, Subeen Park, Sungjun Lim, Yonghan Jung et al.CVPR 2025
- Model Agnostic Sample Reweighting for Out-of-Distribution LearningXiao Zhou, Yong Lin, Renjie Pi, Weizhong Zhang et al.ICML 2022 · 73 citations
- Ensemble Pruning for Out-of-distribution GeneralizationFengchun Qiao, Xi PengICML 2024 · 3 citations
