Convex Calibrated Surrogates for the Multi-Label F-Measure
Mingyuan Zhang, Harish Guruprasad Ramaswamy, Shivani Agarwal
摘要
The F -measure is a widely used performance measure for multi-label classification, where multiple labels can be active in an instance simultaneously (e.g. in image tagging, multiple tags can be active in any image). In particular, the F -measure explicitly balances recall (fraction of active labels predicted to be active) and precision (fraction of labels predicted to be active that are actually so), both of which are important in evaluating the overall performance of a multi-label classifier. As with most discrete prediction problems, however, directly optimizing the F -measure is computationally hard. In this paper, we explore the question of designing convex surrogate losses that are calibrated for the F -measure -specifically, that have the property that minimizing the surrogate loss yields (in the limit of sufficient data) a Bayes optimal multi-label classifier for the F -measure. We show that the F -measure for an s-label problem, when viewed as a 2 s × 2 s loss matrix, has rank at most s 2 + 1, and apply a result of Ramaswamy et al. (2014) to design a family of convex calibrated surrogates for the F -measure. The resulting surrogate risk minimization algorithms can be viewed as decomposing the multi-label F -measure learning problem into s 2 + 1 binary class probability estimation problems. We also provide a quantitative regret transfer bound for our surrogates, which allows any regret guarantees for the binary problems to be transferred to regret guarantees for the overall F -measure problem, and discuss a connection with the algorithm of Dembczynski et al. (2013) . Our experiments confirm our theoretical findings.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper15
- Cross-Entropy Loss Functions: Theoretical Analysis and ApplicationsAnqi Mao, Mehryar Mohri, Yutao ZhongICML 2023 · 被引用 790 次
- Multi-Class -Consistency BoundsPranjal Awasthi, Anqi Mao, Mehryar Mohri, Yutao ZhongNeurIPS 2022 · 被引用 48 次
- Generalizing Consistent Multi-Class Classification with Rejection to be Compatible with Arbitrary LossesYuzhou Cao, Tianchi Cai, Lei Feng, Lihong Gu 等NeurIPS 2022 · 被引用 42 次
- In Defense of Softmax Parametrization for Calibrated and Consistent Learning to DeferYuzhou Cao, Hussein Mozannar, Lei Feng, Hongxin Wei 等NeurIPS 2023 · 被引用 36 次
- H-Consistency Bounds: Characterization and ExtensionsAnqi Mao, Mehryar Mohri, Yutao ZhongNeurIPS 2023 · 被引用 34 次
相关 Paper
- Towards Decision-Friendly AUC: Learning Multi-Classifier with AUCµPeifeng Gao, Qianqian Xu, Peisong Wen, Huiyang Shao 等AAAI 2023 · 被引用 1 次
- Bayes Consistency vs. H-Consistency: The Interplay between Surrogate Loss Functions and the Scoring Function ClassMingyuan Zhang, Shivani AgarwalNeurIPS 2020 · 被引用 42 次
- Multi-label classification: do Hamming loss and subset accuracy really conflict with each other?Guoqiang Wu, Jun ZhuNeurIPS 2020 · 被引用 44 次
- Multi-Label Learning with Stronger Consistency GuaranteesAnqi Mao, Mehryar Mohri, Yutao ZhongNeurIPS 2024 · 被引用 30 次
- Regret Bounds for Multilabel Classification in Sparse Label RegimesRóbert Busa-Fekete, Heejin Choi, Krzysztof Dembczynski, Claudio Gentile 等NeurIPS 2022 · 被引用 5 次
