Continual Origin Tracing of LLM-Generated Text
Haoran Li, Quan Wang
摘要
The rapid development of large language models (LLMs) raises concerns about their potential misuse. Accurately identifying and tracing the origin of LLM-generated content is crucial for accountability and transparency. Previous methods typically frame origin tracing as multi-class classification with a fixed label set, thus struggle to adapt to new LLMs without frequent retraining. This paper introduces a new task, continual origin tracing of LLM-generated text, which frames origin tracing in a continual learning or, more precisely, class-incremental learning manner, where new LLMs continuously emerge, and a model incrementally learns to identify new LLMs without forgetting old ones. A novel training-free method is further devised for the task, which continually extracts prototypes for emerging LLMs using a frozen pre-trained model, and conducts global and local prototype decorrelation to improve prototype matching, thus favoring more accurate tracing. To facilitate evaluation on the new task, we construct a benchmark comprising text generated by 19 recently released LLMs from 12 vendors that simulates a real-world scenario where these LLMs emerge over time and need to be recognized incrementally across 8 diverse domains. Rigorous evaluations on this benchmark highlight the effectiveness and potential of the proposed method in the new task, offering a promising direction for future research.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper1
问问它们各自怎么用它相关 Paper
- Profiler: Black-box AI-generated Text Origin Detection via Context-aware Inference Pattern AnalysisHanxi Guo, Siyuan Cheng, Xiaolong Jin, Zhuo Zhang 等EMNLP 2025
- Investigating How Pre-training Data Leakage Affects Models' Reproduction and Detection CapabilitiesMasahiro Kaneko, Timothy BaldwinEMNLP 2025 · 被引用 2 次
- Matching Pairs: Attributing Fine-Tuned Models to their Pre-Trained Large Language ModelsMyles Foley, Ambrish Rawat, Taesung Lee, Yufang Hou 等ACL 2023 · 被引用 2 次
- M4GT-Bench: Evaluation Benchmark for Black-Box Machine-Generated Text DetectionYuxia Wang, Jonibek Mansurov, Petar Ivanov, Jinyan Su 等ACL 2024
- MAGE: Machine-generated Text Detection in the WildYafu Li, Qintong Li, Leyang Cui, Wei Bi 等ACL 2024 · 被引用 44 次
