Multi-label classification: do Hamming loss and subset accuracy really conflict with each other?
Guoqiang Wu, Jun Zhu
Abstract
Various evaluation measures have been developed for multi-label classification, including Hamming Loss (HL), Subset Accuracy (SA) and Ranking Loss (RL). However, there is a gap between empirical results and the existing theories: 1) an algorithm often empirically performs well on some measure(s) while poorly on others, while a formal theoretical analysis is lacking; and 2) in small label space cases, the algorithms optimizing HL often have comparable or even better performance on the SA measure than those optimizing SA directly, while existing theoretical results show that SA and HL are conflicting measures. This paper provides an attempt to fill up this gap by analyzing the learning guarantees of the corresponding learning algorithms on both SA and HL measures. We show that when a learning algorithm optimizes HL with its surrogate loss, it enjoys an error bound for the HL measure independent of (the number of labels), while the bound for the SA measure depends on at most . On the other hand, when directly optimizing SA with its surrogate loss, it has learning guarantees that depend on for both HL and SA measures. This explains the observation that when the label space is not large, optimizing HL with its surrogate loss can have promising performance for SA. We further show that our techniques are applicable to analyze the learning guarantees of algorithms on other measures, such as RL. Finally, the theoretical analyses are supported by experimental results.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3d32623e-c935-4b0a-9857-f8f357db3f2fCited by top-tier papers7
- Multi-Label Learning with Stronger Consistency GuaranteesAnqi Mao, Mehryar Mohri, Yutao ZhongNeurIPS 2024 · 30 citations
- Rethinking and Reweighting the Univariate Losses for Multi-Label Ranking: Consistency and GeneralizationGuoqiang Wu, Chongxuan Li, Kun Xu, Jun ZhuNeurIPS 2021 · 13 citations
- Fine-grained Generalization Analysis of Vector-Valued LearningLiang Wu, Antoine Ledent, Yunwen Lei, Marius KloftAAAI 2021 · 11 citations
- Generalization Analysis for Multi-Label LearningYifan Zhang, Min-Ling ZhangICML 2024 · 7 citations
- Generalization Analysis for Label-Specific Representation LearningYifan Zhang, Min-Ling ZhangNeurIPS 2024 · 6 citations
Related papers
- Regret Bounds for Multilabel Classification in Sparse Label RegimesRóbert Busa-Fekete, Heejin Choi, Krzysztof Dembczynski, Claudio Gentile et al.NeurIPS 2022 · 5 citations
- Convex Calibrated Surrogates for the Multi-Label F-MeasureMingyuan Zhang, Harish Guruprasad Ramaswamy, Shivani AgarwalICML 2020 · 23 citations
- Reliable Multilabel Classification: Prediction with Partial AbstentionVu-Linh Nguyen, Eyke HüllermeierAAAI 2020 · 17 citations
- Multi-Label Ranking Loss Minimization for Matrix CompletionJiaxuan Li, Xiaoyan Zhu, Hongrui Wang, Yu Zhang et al.AAAI 2025
- Towards Understanding Generalization of Macro-AUC in Multi-label LearningGuoqiang Wu, Chongxuan Li, Yilong YinICML 2023 · 9 citations
