AI Wrote My Paper and All I Got was This False Negative:* Measuring the Efficacy of Commercial AI Text Detectors
Seth Layton, Bernardo B. P. Medeiros, Kevin R. B. Butler, Patrick Traynor
摘要
Academic institutions and publishers are increasingly relying on commercial AI-generated-text (AIGT) detectors to combat plagiarism and verify authorship in the era of large language models (lLMs). As the number of submissions to academic security conferences increases at an exponential rate, the temptation to deploy detectors to protect academic integrity commensurately increases. Unfortunately, even when benchmarks are disclosed, these detectors lack appropriate performance characterizations for use in evaluating academic security writing. In this paper, we conduct a comprehensive empirical evaluation of leading AIGT detector performance on academic security writing. We collect a dataset of all papers () from Tier-1 conferences (IEEE S&P, CCS, NDSS, and USENIX) prior to the public release of ChatGPT. We then create an AIGT version of each of these papers and use this combined dataset to evaluate the top five most popular AIGT detectors, based on Tranco-list rankings. Our evaluation not only finds that performance varies wildly across AIGT detectors (e.g., FPRs between 0.05 % and 68.6 % and FNRs ranging between 0.3 % and 99.6 %). Even more critically, these detectors are trivially circumvented by a simple adaptive adversary (e.g., a TPR reduction from 94.2 % to 2.5 %). Ultimately, the limitations of current detector-based approaches create an adversarial environment in which achieving authorship verification remains out of reach.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- An Empirical Study to Evaluate AIGC Detectors on Code ContentJian Wang, Shangqing Liu, Xiaofei Xie, Yi LiASE 2024 · 被引用 4 次
- Hidding the Ghostwriters: An Adversarial Evaluation of AI-Generated Student Essay DetectionXinlin Peng, Ying Zhou, Ben He, Le Sun 等EMNLP 2023 · 被引用 9 次
- Navigating the Shadows: Unveiling Effective Disturbances for Modern AI Content DetectorsYing Zhou, Ben He, Le SunACL 2024
- How Large Language Models are Transforming Machine-Paraphrase PlagiarismJan Philip Wahle, Terry Ruas, Frederic Kirstein, Bela GippEMNLP 2022 · 被引用 26 次
- Policies Permitting LLM Use for Polishing Peer Reviews Are Currently Not EnforceableRounak Saha, Gurusha Juneja, Dayita Chaudhuri, Naveeja Sajeevan 等ICML 2026 · 被引用 3 次
