Robust Performance Metrics for Authentication Systems
Shridatt Sugrim, Can Liu, Meghan McLean, Janne Lindqvist
摘要
Research has produced many types of authentication systems that use machine learning. However, there is no consistent approach for reporting performance metrics and the reported metrics are inadequate. In this work, we show that several of the common metrics used for reporting performance, such as maximum accuracy (ACC), equal error rate (EER) and area under the ROC curve (AUROC), are inherently flawed. These common metrics hide the details of the inherent tradeoffs a system must make when implemented. Our findings show that current metrics give no insight into how system performance degrades outside the ideal conditions in which they were designed. We argue that adequate performance reporting must be provided to enable meaningful evaluation and that current, commonly used approaches fail in this regard. We present the unnormalized frequency count of scores (FCS) to demonstrate the mathematical underpinnings that lead to these failures and show how they can be avoided. The FCS can be used to augment the performance reporting to enable comparison across systems in a visual way. When reported with the Receiver Operating Characteristics curve (ROC), these two metrics provide a solution to the limitations of currently reported metrics. Finally, we show how to use the FCS and ROC metrics to evaluate and compare different authentication systems.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Inexpensive Brainwave Authentication: New Techniques and Insights on User AcceptancePatricia Arias Cabarcos, Thilo Habrich, Karen Becker, Christian Becker 等USENIX Security 2021 · 被引用 34 次
- "Get in Researchers; We're Measuring Reproducibility": A Reproducibility Study of Machine Learning Papers in Tier 1 Security ConferencesDaniel Olszewski, Allison Lu, Carson Stillman, Kevin Warren 等CCS 2023 · 被引用 19 次
- SoK: The Good, The Bad, and The Unbalanced: Measuring Structural Limitations of Deepfake Media DatasetsSeth Layton, Tyler Tucker, Daniel Olszewski, Kevin Warren 等USENIX Security 2024 · 被引用 11 次
- Dos and Don'ts of Machine Learning in Computer SecurityDaniel Arp, Erwin Quiring, Feargus Pendlebury, Alexander Warnecke 等USENIX Security 2022
- On the Resilience of Biometric Authentication Systems against Random InputsBenjamin Zi Hao Zhao, Hassan Jameel Asghar, Mohamed Ali KâafarNDSS 2020
它引用的顶会 Paper8
- Hearing Your Voice is Not Enough: An Articulatory Gesture Based Liveness Detection for Voice AuthenticationLinghan Zhang, Sheng Tan, Jie YangCCS 2017 · 被引用 212 次
- Who Are You? A Statistical Approach to Measuring User AuthenticityDavid Freeman, Sakshi Jain, Markus Dürmuth, Battista Biggio 等NDSS 2016 · 被引用 151 次
- Multi-touch Authentication Using Hand Geometry and Behavioral InformationYunpeng Song, Zhongmin Cai, Zhi-Li ZhangS&P 2017 · 被引用 98 次
- VibWrite: Towards Finger-input Authentication on Ubiquitous Surfaces via Physical VibrationJian Liu, Chen Wang, Yingying Chen, Nitesh SaxenaCCS 2017 · 被引用 93 次
- Using Reflexive Eye Movements for Fast Challenge-Response AuthenticationIvo Sluganovic, Marc Roeschlin, Kasper Bonne Rasmussen, Ivan MartinovicCCS 2016 · 被引用 93 次
相关 Paper
- A Closer Look at AUROC and AUPRC under Class ImbalanceMatthew B. A. McDermott, Haoran Zhang, Lasse Hyldig Hansen, Giovanni Angelotti 等NeurIPS 2024 · 被引用 191 次
- Overcoming Common Flaws in the Evaluation of Selective Classification SystemsJeremias Traub, Till J. Bungert, Carsten T. Lüth, Michael Baumgartner 等NeurIPS 2024 · 被引用 44 次
- Never mind the metrics - what about the uncertainty? Visualising binary confusion matrix metric distributions to put performance in perspectiveDavid R. Lovell, Dimity Miller, Jaiden Capra, Andrew P. BradleyICML 2023 · 被引用 3 次
- The VOROS: Lifting ROC Curves to 3D to Summarize Unbalanced Classifier PerformanceChristopher Ratigan, Lenore CowenAAAI 2025 · 被引用 1 次
- Asymptotically Unbiased Instance-wise Regularized Partial AUC Optimization: Theory and AlgorithmHuiyang Shao, Qianqian Xu, Zhiyong Yang, Shilong Bao 等NeurIPS 2022 · 被引用 7 次
