Boosted CVaR Classification
Runtian Zhai, Chen Dan, Arun Sai Suggala, J. Zico Kolter, Pradeep Ravikumar
Abstract
Many modern machine learning tasks require models with high tail performance, i.e. high performance over the worst-off samples in the dataset. This problem has been widely studied in fields such as algorithmic fairness, class imbalance, and risk-sensitive decision making. A popular approach to maximize the model's tail performance is to minimize the CVaR (Conditional Value at Risk) loss, which computes the average risk over the tails of the loss. However, for classification tasks where models are evaluated by the zero-one loss, we show that if the classifiers are deterministic, then the minimizer of the average zero-one loss also minimizes the CVaR zero-one loss, suggesting that CVaR loss minimization is not helpful without additional assumptions. We circumvent this negative result by minimizing the CVaR loss over randomized classifiers, for which the minimizers of the average zero-one loss and the CVaR zero-one loss are no longer the same, so minimizing the latter can lead to better tail performance. To learn such randomized classifiers, we propose the Boosted CVaR Classification framework which is motivated by a direct relationship between CVaR and a classical boosting algorithm called LPBoost. Based on this framework, we design an algorithm called -AdaLPBoost. We empirically evaluate our proposed algorithm on four benchmark datasets and show that it achieves higher tail performance than deterministic model training methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3fc9e529-9ff3-4dec-99cf-af2fadcef935Cited by top-tier papers7
- UMIX: Improving Importance Weighting for Subpopulation Shift via Uncertainty-Aware MixupZongbo Han, Zhipeng Liang, Fan Yang, Liu Liu et al.NeurIPS 2022 · 53 citations
- Distributionally Robust Optimization via Ball Oracle AccelerationYair Carmon, Danielle HauslerNeurIPS 2022 · 23 citations
- Understanding Why Generalized Reweighting Does Not Improve Over ERMRuntian Zhai, Chen Dan, J. Zico Kolter, Pradeep Kumar RavikumarICLR 2023 · 6 citations
- Prompting is a Double-Edged Sword: Improving Worst-Group Robustness of Foundation ModelsAmrith Setlur, Saurabh Garg, Virginia Smith, Sergey LevineICML 2024 · 4 citations
- Criterion Collapse and Loss Distribution ControlMatthew J. HollandICML 2024 · 2 citations
Builds on6
- An Investigation of Why Overparameterization Exacerbates Spurious CorrelationsShiori Sagawa, Aditi Raghunathan, Pang Wei Koh, Percy LiangICML 2020 · 436 citations
- Fairness without Demographics through Adversarially Reweighted LearningPreethi Lahoti, Alex Beutel, Jilin Chen, Kang Lee et al.NeurIPS 2020 · 406 citations
- DORO: Distributional and Outlier Robust OptimizationRuntian Zhai, Chen Dan, J. Zico Kolter, Pradeep RavikumarICML 2021 · 74 citations
- Class-Weighted Classification: Trade-offs and Robust ApproachesZiyu Xu, Chen Dan, Justin Khim, Pradeep RavikumarICML 2020 · 53 citations
- Modeling the Second Player in Distributionally Robust OptimizationPaul Michel, Tatsunori Hashimoto, Graham NeubigICLR 2021 · 39 citations
Related papers
- Safe Collaborative FilteringRiku Togashi, Tatsushi Oka, Naoto Ohsaka, Tetsuro MorimuraICLR 2024 · 2 citations
- Regret Bounds for Risk-Sensitive Reinforcement LearningOsbert Bastani, Yecheng Jason Ma, Estelle Shen, Wanqiao XuNeurIPS 2022 · 29 citations
- Fair Wrapping for Black-box PredictionsAlexander Soen, Ibrahim M. Alabdulmohsin, Sanmi Koyejo, Yishay Mansour et al.NeurIPS 2022 · 8 citations
- PAC-Bayesian Bound for the Conditional Value at RiskZakaria Mhammedi, Benjamin Guedj, Robert C. WilliamsonNeurIPS 2020 · 25 citations
- Adaptive Sampling for Stochastic Risk-Averse LearningSebastian Curi, Kfir Y. Levy, Stefanie Jegelka, Andreas KrauseNeurIPS 2020 · 65 citations
