Principled Algorithms for Optimizing Generalized Metrics in Binary Classification
Anqi Mao, Mehryar Mohri, Yutao Zhong
Abstract
In applications with significant class imbalance or asymmetric costs, metrics such as the F βmeasure, AM measure, Jaccard similarity coefficient, and weighted accuracy offer more suitable evaluation criteria than standard binary classification loss. However, optimizing these metrics present significant computational and statistical challenges. Existing approaches often rely on the characterization of the Bayes-optimal classifier, and use threshold-based methods that first estimate class probabilities and then seek an optimal threshold. This leads to algorithms that are not tailored to restricted hypothesis sets and lack finite-sample performance guarantees. In this work, we introduce principled algorithms for optimizing generalized metrics, supported by H-consistency and finite-sample generalization bounds. Our approach reformulates metric optimization as a generalized cost-sensitive learning problem, enabling the design of novel surrogate loss functions with provable H-consistency guarantees. Leveraging this framework, we develop new algorithms, METRO (Metric Optimization), with strong theoretical performance guarantees. We report the results of experiments demonstrating the effectiveness of our methods compared to prior baselines.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c30cb09c-ceb8-4f7d-b2c3-1c07e5bb833bCited by top-tier papers9
- Improved Balanced Classification with Theoretically Grounded Loss FunctionsCorinna Cortes, Mehryar Mohri, Yutao ZhongNeurIPS 2025 · 19 citations
- Why Ask One When You Can Ask k? Learning-to-Defer to the Top-k ExpertsYannis Montreuil, Axel Carlier, Lai Xing Ng, Wei Tsang OoiICLR 2026 · 7 citations
- Linear-Core Surrogates: Smooth Loss Functions with Linear Rates for Classification and Structured PredictionMehryar Mohri, Yutao ZhongICML 2026 · 7 citations
- A Theoretical Framework for Modular Learning of Robust Generative ModelsCorinna Cortes, Mehryar Mohri, Yutao ZhongICML 2026 · 7 citations
- Optimized Deferral for Imbalanced SettingsCorinna Cortes, Anqi Mao, Mehryar Mohri, Yutao ZhongICML 2026 · 7 citations
Builds on29
- Cross-Entropy Loss Functions: Theoretical Analysis and ApplicationsAnqi Mao, Mehryar Mohri, Yutao ZhongICML 2023 · 790 citations
- Two-Stage Learning to Defer with Multiple ExpertsAnqi Mao, Christopher Mohri, Mehryar Mohri, Yutao ZhongNeurIPS 2023 · 98 citations
- Calibration and Consistency of Adversarial Surrogate LossesPranjal Awasthi, Natalie Frank, Anqi Mao, Mehryar Mohri et al.NeurIPS 2021 · 59 citations
- H-Consistency Bounds for Surrogate Loss MinimizersPranjal Awasthi, Anqi Mao, Mehryar Mohri, Yutao ZhongICML 2022 · 50 citations
- Multi-Class -Consistency BoundsPranjal Awasthi, Anqi Mao, Mehryar Mohri, Yutao ZhongNeurIPS 2022 · 48 citations
Related papers
- Balancing the Scales: A Theoretical and Algorithmic Framework for Learning from Imbalanced DataCorinna Cortes, Anqi Mao, Mehryar Mohri, Yutao ZhongICML 2025
- Towards Decision-Friendly AUC: Learning Multi-Classifier with AUCµPeifeng Gao, Qianqian Xu, Peisong Wen, Huiyang Shao et al.AAAI 2023 · 1 citation
- Convex Calibrated Surrogates for the Multi-Label F-MeasureMingyuan Zhang, Harish Guruprasad Ramaswamy, Shivani AgarwalICML 2020 · 23 citations
- Aligning Evaluation with Clinical Priorities: Calibration, Label Shift, and Error CostsGerardo Flores, Alyssa H. Smith, Julia Fukuyama, Ashia C. WilsonNeurIPS 2025 · 5 citations
- Multi-Label Learning with Stronger Consistency GuaranteesAnqi Mao, Mehryar Mohri, Yutao ZhongNeurIPS 2024 · 30 citations
