Modeling the Second Player in Distributionally Robust Optimization
Paul Michel, Tatsunori Hashimoto, Graham Neubig
摘要
Distributionally robust optimization (DRO) provides a framework for training machine learning models that are able to perform well on a collection of related data distributions (the "uncertainty set"). This is done by solving a min-max game: the model is trained to minimize its maximum expected loss among all distributions in the uncertainty set. While careful design of the uncertainty set is critical to the success of the DRO procedure, previous work has been limited to relatively simple alternatives that keep the min-max optimization problem exactly tractable, such as f -divergence balls. In this paper, we argue instead for the use of neural generative models to characterize the worst-case distribution, allowing for more flexible and problem-specific selection of the uncertainty set. However, while simple conceptually, this approach poses a number of implementation and optimization challenges. To circumvent these issues, we propose a relaxation of the KL-constrained inner maximization objective that makes the DRO problem more amenable to gradient-based optimization of large scale generative models, and develop model selection heuristics to guide hyper-parameter search. On both toy settings and realistic NLP tasks, we find that the proposed approach yields models that are more robust than comparable baselines 1 .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper21
- Examining and Combating Spurious Features under Distribution ShiftChunting Zhou, Xuezhe Ma, Paul Michel, Graham NeubigICML 2021 · 被引用 78 次
- DORO: Distributional and Outlier Robust OptimizationRuntian Zhai, Chen Dan, J. Zico Kolter, Pradeep RavikumarICML 2021 · 被引用 74 次
- Learning to Augment Distributions for Out-of-distribution DetectionQizhou Wang, Zhen Fang, Yonggang Zhang, Feng Liu 等NeurIPS 2023 · 被引用 59 次
- Understanding Contrastive Learning via Distributionally Robust OptimizationJunkang Wu, Jiawei Chen, Jiancan Wu, Wentao Shi 等NeurIPS 2023 · 被引用 55 次
- UMIX: Improving Importance Weighting for Subpopulation Shift via Uncertainty-Aware MixupZongbo Han, Zhipeng Liang, Fan Yang, Liu Liu 等NeurIPS 2022 · 被引用 53 次
它引用的顶会 Paper5
- Don't Stop Pretraining: Adapt Language Models to Domains and TasksSuchin Gururangan, Ana Marasovic, Swabha Swayamdipta, Kyle Lo 等ACL 2020 · 被引用 93 次
- Distributionally Robust Counterfactual Risk MinimizationLouis Faury, Ugo Tanielian, Elvis Dohmatob, Elena Smirnova 等AAAI 2020 · 被引用 48 次
- Distributional Robustness with IPMs and links to Regularization and GANsHisham HusainNeurIPS 2020 · 被引用 25 次
- Robust Bayesian Classification Using An Optimistic Score RatioViet Anh Nguyen, Nian Si, Jose H. BlanchetICML 2020 · 被引用 15 次
- Towards Debiasing NLU Models from Unknown BiasesPrasetya Ajie Utama, Nafise Sadat Moosavi, Iryna GurevychEMNLP 2020 · 被引用 3 次
相关 Paper
- Distributionally Robust Optimization via Generative Ambiguity ModelingJiaqi Wen, Jianyi YangICLR 2026 · 被引用 3 次
- Distributionally Robust Models with Parametric Likelihood RatiosPaul Michel, Tatsunori Hashimoto, Graham NeubigICLR 2022 · 被引用 21 次
- An Online Method for A Class of Distributionally Robust Optimization with Non-convex ObjectivesQi Qi, Zhishuai Guo, Yi Xu, Rong Jin 等NeurIPS 2021 · 被引用 61 次
- Generalization Bounds with Minimal Dependency on Hypothesis Class via Distributionally Robust OptimizationYibo Zeng, Henry LamNeurIPS 2022 · 被引用 11 次
- Outlier-Robust Distributionally Robust Optimization via Unbalanced Optimal TransportZifan Wang, Yi Shen, Michael M. Zavlanos, Karl Henrik JohanssonNeurIPS 2024 · 被引用 16 次
