AutoBalance: Optimized Loss Functions for Imbalanced Data
Mingchen Li, Xuechen Zhang, Christos Thrampoulidis, Jiasi Chen, Samet Oymak
摘要
Imbalanced datasets are commonplace in modern machine learning problems. The presence of under-represented classes or groups with sensitive attributes results in concerns about generalization and fairness. Such concerns are further exacerbated by the fact that large capacity deep nets can perfectly fit the training data and appear to achieve perfect accuracy and fairness during training, but perform poorly during test. To address these challenges, we propose AutoBalance, a bi-level optimization framework that automatically designs a training loss function to optimize a blend of accuracy and fairness-seeking objectives. Specifically, a lower-level problem trains the model weights, and an upper-level problem tunes the loss function by monitoring and optimizing the desired objective over the validation data. Our loss design enables personalized treatment for classes/groups by employing a parametric cross-entropy loss and individualized data augmentation schemes. We evaluate the benefits and performance of our approach for the application scenarios of imbalanced and group-sensitive classification. Extensive empirical evaluations demonstrate the benefits of AutoBalance over state-of-the-art approaches. Our experimental findings are complemented with theoretical insights on loss function design and the benefits of train-validation split. All code is available open-source.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper18
- Imbalance Trouble: Revisiting Neural-Collapse GeometryChristos Thrampoulidis, Ganesh Ramachandra Kini, Vala Vakilian, Tina BehniaNeurIPS 2022 · 被引用 101 次
- FedNest: Federated Bilevel, Minimax, and Compositional OptimizationDavoud Ataee Tarzanagh, Mingchen Li, Christos Thrampoulidis, Samet OymakICML 2022 · 被引用 85 次
- SoLar: Sinkhorn Label Refinery for Imbalanced Partial-Label LearningHaobo Wang, Mingxuan Xia, Yixuan Li, Yuren Mao 等NeurIPS 2022 · 被引用 54 次
- Selective Attention: Enhancing Transformer through Principled Context ControlXuechen Zhang, Xiangyu Chang, Mingchen Li, Amit K. Roy-Chowdhury 等NeurIPS 2024 · 被引用 32 次
- Pure Noise to the Rescue of Insufficient Data: Improving Imbalanced Classification by Training on Random Noise ImagesShiran Zada, Itay Benou, Michal IraniICML 2022 · 被引用 31 次
它引用的顶会 Paper11
- CutMix: Regularization Strategy to Train Strong Classifiers With Localizable FeaturesSangdoo Yun, Dongyoon Han, Sanghyuk Chun, Seong Joon Oh 等ICCV 2019 · 被引用 5,843 次
- Decoupling Representation and Classifier for Long-Tailed RecognitionBingyi Kang, Saining Xie, Marcus Rohrbach, Zhicheng Yan 等ICLR 2020 · 被引用 1,496 次
- Long-tail learning via logit adjustmentAditya Krishna Menon, Sadeep Jayasumana, Ankit Singh Rawat, Himanshu Jain 等ICLR 2021 · 被引用 937 次
- A Group-Theoretic Framework for Data AugmentationShuxiao Chen, Edgar Dobriban, Jane H. LeeNeurIPS 2020 · 被引用 254 次
- Polylogarithmic width suffices for gradient descent to achieve arbitrarily small test error with shallow ReLU networksZiwei Ji, Matus TelgarskyICLR 2020 · 被引用 193 次
相关 Paper
- Fair Bilevel Neural Network (FairBiNN): On Balancing fairness and accuracy via Stackelberg EquilibriumMehdi Yazdani-Jahromi, Ali Khodabandeh Yalabadi, Amirarsalan Rajabi, Aida Tayebi 等NeurIPS 2024 · 被引用 11 次
- Class-Attribute Priors: Adapting Optimization to Heterogeneity and Fairness ObjectiveXuechen Zhang, Mingchen Li, Jiasi Chen, Christos Thrampoulidis 等AAAI 2024 · 被引用 3 次
- A Unified Loss for Handling Inter-Class and Intra-Class Imbalance in Medical Image SegmentationFeilong Xu, Feiyang Yang, Xiongfei Li, Xiaoli ZhangAAAI 2025 · 被引用 6 次
- Re-weighting Based Group Fairness Regularization via Classwise Robust OptimizationSangwon Jung, Taeeon Park, Sanghyuk Chun, Taesup MoonICLR 2023 · 被引用 5 次
- FIFA: Making Fairness More Generalizable in Classifiers Trained on Imbalanced DataZhun Deng, Jiayao Zhang, Linjun Zhang, Ting Ye 等ICLR 2023 · 被引用 7 次
