Label Distributionally Robust Losses for Multi-class Classification: Consistency, Robustness and Adaptivity
Dixian Zhu, Yiming Ying, Tianbao Yang
摘要
We study a family of loss functions named label-distributionally robust (LDR) losses for multi-class classification that are formulated from distributionally robust optimization (DRO) perspective, where the uncertainty in the given label information are modeled and captured by taking the worse case of distributional weights. The benefits of this perspective are several fold: (i) it provides a unified framework to explain the classical cross-entropy (CE) loss and SVM loss and their variants, (ii) it includes a special family corresponding to the temperature-scaled CE loss, which is widely adopted but poorly understood; (iii) it allows us to achieve adaptivity to the uncertainty degree of label information at an instance level. Our contributions include: (1) we study both consistency and robustness by establishing top- () consistency of LDR losses for multi-class classification, and a negative result that a top- consistent and symmetric robust loss cannot achieve top- consistency simultaneously for all ; (2) we propose a new adaptive LDR loss that automatically adapts the individualized temperature parameter to the noise degree of class label of each instance; (3) we demonstrate stable and competitive performance for the proposed adaptive LDR loss on 7 benchmark datasets under 6 noisy label and 1 clean settings against 13 loss functions, and on one real-world noisy dataset. The code is open-sourced at https://github.com/Optimization-AI/ICML2023_LDR.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- To Cool or not to Cool? Temperature Network Meets Large Foundation Models via DROZi-Hao Qiu, Siqi Guo, Mao Xu, Tuo Zhao 等ICML 2024 · 被引用 11 次
- -Softmax: Approximating One-Hot Vectors for Mitigating Label NoiseJialiang Wang, Xiong Zhou, Deming Zhai, Junjun Jiang 等NeurIPS 2024 · 被引用 10 次
- Blockwise Stochastic Variance-Reduced Methods with Parallel Speedup for Multi-Block Bilevel OptimizationQuanqi Hu, Zi-Hao Qiu, Zhishuai Guo, Lijun Zhang 等ICML 2023 · 被引用 9 次
- Statistical Consistency and Generalization of Contrastive Representation LearningYuanfan Li, Xiyuan Wei, Tianbao Yang, Yiming YingICML 2026 · 被引用 1 次
- Gradient Aligned Regression via Pairwise LossesDixian Zhu, Tianbao Yang, Livnat JerbyICML 2025
它引用的顶会 Paper10
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- Normalized Loss Functions for Deep Learning with Noisy LabelsXingjun Ma, Hanxun Huang, Yisen Wang, Simone Romano 等ICML 2020 · 被引用 547 次
- Large-Scale Methods for Distributionally Robust OptimizationDaniel Levy, Yair Carmon, John C. Duchi, Aaron SidfordNeurIPS 2020 · 被引用 281 次
- Generalized Jensen-Shannon Divergence Loss for Learning with Noisy LabelsErik Englesson, Hossein AzizpourNeurIPS 2021 · 被引用 170 次
相关 Paper
- Wasserstein Distributional Normalization For Robust Distributional Certification of Noisy Labeled DataSung Woo Park, Junseok KwonICML 2021 · 被引用 4 次
- Probability Guided Loss for Long-Tailed Multi-Label Image ClassificationDekun LinAAAI 2023 · 被引用 17 次
- DRAUC: An Instance-wise Distributionally Robust AUC Optimization FrameworkSiran Dai, Qianqian Xu, Zhiyong Yang, Xiaochun Cao 等NeurIPS 2023 · 被引用 5 次
- Advancing Loss Functions in Recommender Systems: A Comparative Study with a Rényi Divergence-Based SolutionShengjia Zhang, Jiawei Chen, Changdong Li, Sheng Zhou 等AAAI 2025 · 被引用 6 次
- Coping with Label Shift via Distributionally Robust OptimisationJingzhao Zhang, Aditya Krishna Menon, Andreas Veit, Srinadh Bhojanapalli 等ICLR 2021 · 被引用 79 次
