Who Wrote This Line? Evaluating the Detection of LLM-Generated Classical Chinese Poetry
Jiang Li, Tian Lan, Shanshan Wang, Zdongxing, Dianqing Lin, Guanglai Gao, Derek F. Wong, Xiangdong Su
摘要
The rapid development of large language models (LLMs) has extended text generation tasks into the literary domain. However, AIgenerated literary creations has raised increasingly prominent issues of creative authenticity and ethics in literary world, making the detection of LLM-generated literary texts essential and urgent. While previous works have made significant progress in detecting AI-generated text, it has yet to address classical Chinese poetry. Due to the unique linguistic features of classical Chinese poetry, such as strict metrical regularity, a shared system of poetic imagery, and flexible syntax, distinguishing whether a poem is authored by AI presents a substantial challenge. To address these issues, we introduce ChangAn, a benchmark for detecting LLM-generated classical Chinese poetry that containing total 30,664 poems, 10,276 are human-written poems and 20,388 poems are generated by four popular LLMs. Based on ChangAn, we conducted a systematic evaluation of 12 AI detectors, investigating their performance variations across different text granularities and generation strategies. Our findings highlight the limitations of current Chinese text detectors, which fail to serve as reliable tools for detecting LLM-generated classical Chinese poetry. These results validate the effectiveness and necessity of our proposed ChangAn benchmark. Our dataset and code are available at https://github.com/VelikayaScarlet/ ChangAn . 1 AI诗歌成为获奖"钉子户","反AI诗歌联盟群"群主 希望相关部门对AI作品投稿尽快出台相关政策 2 《诗刊》副主编对AI诗歌投稿发出警告 "谁在写"引 发文学圈深度探讨|封面头条 3 用AI诗作投稿:不是走捷径,而是绕远路
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper10
- LLM Evaluators Recognize and Favor Their Own GenerationsArjun Panickssery, Samuel R. Bowman, Shi FengNeurIPS 2024 · 被引用 865 次
- Fast-DetectGPT: Efficient Zero-Shot Detection of Machine-Generated Text via Conditional Probability CurvatureGuangsheng Bao, Yanbin Zhao, Zhiyang Teng, Linyi Yang 等ICLR 2024 · 被引用 311 次
- Multiscale Positive-Unlabeled Detection of AI-Generated TextsYuchuan Tian, Hanting Chen, Xutao Wang, Zheyuan Bai 等ICLR 2024 · 被引用 84 次
- Failure-Driven Workflow RefinementJusheng Zhang, Jing Yang, Kaitong Cai, Ziliang Chen 等ICML 2026 · 被引用 17 次
- Between Lines of Code: Unraveling the Distinct Patterns of Machine and Human ProgrammersYuling Shi, Hongyu Zhang, Chengcheng Wan, Xiaodong GuICSE 2025 · 被引用 9 次
相关 Paper
- Neo-Classic: A Benchmark for Evaluating Linguistic-Aesthetic Reasoning in Classical Chinese PoetryHan Zhang, Zihan Gu, Zhiyuan Wang, Tianyi Ma 等ACL 2026
- MAGE: Machine-generated Text Detection in the WildYafu Li, Qintong Li, Leyang Cui, Wei Bi 等ACL 2024 · 被引用 44 次
- PoeTone: A Framework for Constrained Generation of Structured Chinese Songci with LLMsZhan Qu, Shuzhou Yuan, Michael FärberAAAI 2026 · 被引用 2 次
- POEMetric: The Last Stanza of HumanityBingru Li, Han Wang, Hazel WilkinsonICLR 2026 · 被引用 2 次
- Is Your Paper Being Reviewed by an LLM? Benchmarking AI Text Detection in Peer ReviewSungduk Yu, Man Luo, Avinash Madasu, Vasudev Lal 等ICLR 2026 · 被引用 24 次
