Training Over-parameterized Models with Non-decomposable Objectives
Harikrishna Narasimhan, Aditya Krishna Menon
摘要
Many modern machine learning applications come with complex and nuanced design goals such as minimizing the worst-case error, satisfying a given precision or recall target, or enforcing group-fairness constraints. Popular techniques for optimizing such non-decomposable objectives reduce the problem into a sequence of cost-sensitive learning tasks, each of which is then solved by re-weighting the training loss with example-specific costs. We point out that the standard approach of re-weighting the loss to incorporate label costs can produce unsatisfactory results when used to train over-parameterized models. As a remedy, we propose new cost-sensitive losses that extend the classical idea of logit adjustment to handle more general cost matrices. Our losses are calibrated, and can be further improved with distilled labels from a teacher model. Through experiments on benchmark image datasets, we showcase the effectiveness of our approach in training ResNet models with common robust and constrained optimization objectives.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Long-Tailed Partial Label Learning via Dynamic RebalancingFeng Hong, Jiangchao Yao, Zhihan Zhou, Ya Zhang 等ICLR 2023 · 被引用 8 次
- Learning to Reject Meets Long-tail LearningHarikrishna Narasimhan, Aditya Krishna Menon, Wittawat Jitkrittum, Neha Gupta 等ICLR 2024 · 被引用 7 次
- Cost-Sensitive Self-Training for Optimizing Non-Decomposable MetricsHarsh Rangwani, Shrinivas Ramasubramanian, Sho Takemori, Kato Takashi 等NeurIPS 2022 · 被引用 7 次
- FedFACT: A Provable Framework for Controllable Group-Fairness Calibration in Federated LearningLi Zhang, Zhongxuan Han, Xiaohua Feng, Jiaming Zhang 等NeurIPS 2025 · 被引用 2 次
- Selective Mixup Fine-Tuning for Optimizing Non-Decomposable ObjectivesShrinivas Ramasubramanian, Harsh Rangwani, Sho Takemori, Kunal Samanta 等ICLR 2024 · 被引用 2 次
它引用的顶会 Paper16
- Distributionally Robust Neural NetworksShiori Sagawa, Pang Wei Koh, Tatsunori B. Hashimoto, Percy LiangICLR 2020 · 被引用 1,578 次
- Decoupling Representation and Classifier for Long-Tailed RecognitionBingyi Kang, Saining Xie, Marcus Rohrbach, Zhicheng Yan 等ICLR 2020 · 被引用 1,496 次
- Deep Double Descent: Where Bigger Models and More Data HurtPreetum Nakkiran, Gal Kaplun, Yamini Bansal, Tristan Yang 等ICLR 2020 · 被引用 1,108 次
- Long-tail learning via logit adjustmentAditya Krishna Menon, Sadeep Jayasumana, Ankit Singh Rawat, Himanshu Jain 等ICLR 2021 · 被引用 937 次
- Balanced Meta-Softmax for Long-Tailed Visual RecognitionJiawei Ren, Cunjun Yu, Shunan Sheng, Xiao Ma 等NeurIPS 2020 · 被引用 861 次
相关 Paper
- Distributionally Robust Post-hoc Classifiers under Prior ShiftsJiaheng Wei, Harikrishna Narasimhan, Ehsan Amid, Wen-Sheng Chu 等ICLR 2023
- Fairness without Demographics through Knowledge DistillationJunyi Chai, Taeuk Jang, Xiaoqian WangNeurIPS 2022 · 被引用 57 次
- Re-weighting Based Group Fairness Regularization via Classwise Robust OptimizationSangwon Jung, Taeeon Park, Sanghyuk Chun, Taesup MoonICLR 2023 · 被引用 5 次
- Label-Imbalanced and Group-Sensitive Classification under OverparameterizationGanesh Ramachandra Kini, Orestis Paraskevas, Samet Oymak, Christos ThrampoulidisNeurIPS 2021 · 被引用 122 次
- A Unified Generalization Analysis of Re-Weighting and Logit-Adjustment for Imbalanced LearningZitai Wang, Qianqian Xu, Zhiyong Yang, Yuan He 等NeurIPS 2023 · 被引用 15 次
