DORO: Distributional and Outlier Robust Optimization
Runtian Zhai, Chen Dan, J. Zico Kolter, Pradeep Ravikumar
摘要
Many machine learning tasks involve subpopulation shift where the testing data distribution is a subpopulation of the training distribution. For such settings, a line of recent work has proposed the use of a variant of empirical risk minimization(ERM) known as distributionally robust optimization (DRO). In this work, we apply DRO to real, large-scale tasks with subpopulation shift, and observe that DRO performs relatively poorly, and moreover has severe instability. We identify one direct cause of this phenomenon: sensitivity of DRO to outliers in the datasets. To resolve this issue, we propose the framework of DORO, for Distributional and Outlier Robust Optimization. At the core of this approach is a refined risk function which prevents DRO from overfitting to potential outliers. We instantiate DORO for the Cressie-Read family of Rényi divergence, and delve into two specific instances of this family: CVaR and χ 2 -DRO. We theoretically prove the effectiveness of the proposed method, and empirically show that DORO improves the performance and stability of DRO with experiments on large modern datasets, thereby positively addressing the open question raised by (Hashimoto et al., 2018) . Codes are available at https://github.com/RuntianZ/doro .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper27
- On the Theories Behind Hard Negative Sampling for RecommendationWentao Shi, Jiawei Chen, Fuli Feng, Jizhi Zhang 等WWW 2023 · 被引用 66 次
- Understanding Contrastive Learning via Distributionally Robust OptimizationJunkang Wu, Jiawei Chen, Jiancan Wu, Wentao Shi 等NeurIPS 2023 · 被引用 55 次
- UMIX: Improving Importance Weighting for Subpopulation Shift via Uncertainty-Aware MixupZongbo Han, Zhipeng Liang, Fan Yang, Liu Liu 等NeurIPS 2022 · 被引用 53 次
- Distributionally Robust Graph-based Recommendation SystemBohao Wang, Jiawei Chen, Changdong Li, Sheng Zhou 等WWW 2024 · 被引用 42 次
- Subgroup Robustness Grows On Trees: An Empirical Baseline InvestigationJosh Gardner, Zoran Popovic, Ludwig SchmidtNeurIPS 2022 · 被引用 27 次
它引用的顶会 Paper6
- In Search of Lost Domain GeneralizationIshaan Gulrajani, David Lopez-PazICLR 2021 · 被引用 1,416 次
- An Investigation of Why Overparameterization Exacerbates Spurious CorrelationsShiori Sagawa, Aditi Raghunathan, Pang Wei Koh, Percy LiangICML 2020 · 被引用 436 次
- Fairness without Demographics through Adversarially Reweighted LearningPreethi Lahoti, Alex Beutel, Jilin Chen, Kang Lee 等NeurIPS 2020 · 被引用 406 次
- Class-Weighted Classification: Trade-offs and Robust ApproachesZiyu Xu, Chen Dan, Justin Khim, Pradeep RavikumarICML 2020 · 被引用 53 次
- Learning Bounds for Risk-sensitive LearningJaeho Lee, Sejun Park, Jinwoo ShinNeurIPS 2020 · 被引用 52 次
相关 Paper
- Large-Scale Non-convex Stochastic Constrained Distributionally Robust OptimizationQi Zhang, Yi Zhou, Ashley Prater-Bennette, Lixin Shen 等AAAI 2024 · 被引用 6 次
- Multi-Expert Distributionally Robust Optimization for Out-of-Distribution GeneralizationJinyong Jeong, Hyungu Kahng, Seoung Bum KimNeurIPS 2025 · 被引用 6 次
- Understanding Why Generalized Reweighting Does Not Improve Over ERMRuntian Zhai, Chen Dan, J. Zico Kolter, Pradeep Kumar RavikumarICLR 2023 · 被引用 6 次
- Distributionally Robust Optimization with Data GeometryJiashuo Liu, Jiayun Wu, Bo Li, Peng CuiNeurIPS 2022 · 被引用 28 次
- Model Agnostic Sample Reweighting for Out-of-Distribution LearningXiao Zhou, Yong Lin, Renjie Pi, Weizhong Zhang 等ICML 2022 · 被引用 73 次
