Is a Large Language Model a Good Annotator for Event Extraction?
Ruirui Chen, Chengwei Qin, Weifeng Jiang, Dongkyu Choi
Abstract
Event extraction is an important task in natural language processing that focuses on mining event-related information from unstructured text. Despite considerable advancements, it is still challenging to achieve satisfactory performance in this task, and issues like data scarcity and imbalance obstruct progress. In this paper, we introduce an innovative approach where we employ Large Language Models (LLMs) as expert annotators for event extraction. We strategically include sample data from the training dataset in the prompt as a reference, ensuring alignment between the data distribution of LLM-generated samples and that of the benchmark dataset. This enables us to craft an augmented dataset that complements existing benchmarks, alleviating the challenges of data imbalance and scarcity and thereby enhancing the performance of fine-tuned models. We conducted extensive experiments to validate the efficacy of our proposed method, and we believe that this approach holds great potential for propelling the development and application of more advanced and reliable event extraction systems in real-world scenarios.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d8017f95-a739-43bf-8e40-11429b6eba27Cited by top-tier papers12
- Are LLMs Better than Reported? Detecting Label Errors and Mitigating Their Effect on Model PerformanceOmer Nahum, Nitay Calderon, Orgad Keller, Idan Szpektor et al.EMNLP 2025 · 9 citations
- ACT as Human: Multimodal Large Language Model Data Annotation with Critical ThinkingLequan Lin, Dai Shi, Andi Han, Feng Chen et al.NeurIPS 2025 · 5 citations
- SEOE: A Scalable and Reliable Semantic Evaluation Framework for Open Domain Event DetectionYi-Fan Lu, Xian-Ling Mao, Tian Lan, Tong Zhang et al.ACL 2025 · 3 citations
- Human-Human-AI Triadic Programming: Uncovering the Role of AI Agent and the Value of Human Partner in Collaborative LearningTaufiq Daryanto, Xiaohan Ding, Kaike Ping, Lance T. Wilhelm et al.CHI 2026 · 2 citations
- DiCoRe: Enhancing Zero-shot Event Detection via Divergent-Convergent LLM ReasoningTanmay Parekh, Kartik Mehta, Ninareh Mehrabi, Kai-Wei Chang et al.EMNLP 2025 · 1 citation
Builds on9
- Large Language Models are Zero-Shot ReasonersTakeshi Kojima, Shixiang Shane Gu, Machel Reid, Yutaka Matsuo et al.NeurIPS 2022 · 8,168 citations
- PaLM-E: An Embodied Multimodal Language ModelDanny Driess, Fei Xia, Mehdi S. M. Sajjadi, Corey Lynch et al.ICML 2023 · 2,601 citations
- Is ChatGPT a General-Purpose Natural Language Processing Task Solver?Chengwei Qin, Aston Zhang, Zhuosheng Zhang, Jiaao Chen et al.EMNLP 2023 · 449 citations
- Event Extraction by Answering (Almost) Natural QuestionsXinya Du, Claire CardieEMNLP 2020 · 391 citations
- Event Extraction as Machine Reading ComprehensionJian Liu, Yubo Chen, Kang Liu, Wei Bi et al.EMNLP 2020 · 300 citations
Related papers
- STAR: Boosting Low-Resource Information Extraction by Structure-to-Text Data Generation with Large Language ModelsMingyu Derek Ma, Xiaoxuan Wang, Po-Nien Kung, P. Jeffrey Brantingham et al.AAAI 2024 · 22 citations
- Reflective Agreement: Combining Self-Mixture of Agents with a Sequence Tagger for Robust Event ExtractionFatemeh Haji, Mazal Bethany, Cho-Yu Jason Chiang, Anthony Rios et al.EMNLP 2025
- Towards Robust Universal Information Extraction: Dataset, Evaluation, and SolutionJizhao Zhu, Akang Shi, Zixuan Li, Long Bai et al.ACL 2025
- Generative Multimodal Data Augmentation for Low-Resource Multimodal Named Entity RecognitionZiyan Li, Jianfei Yu, Jia Yang, Wenya Wang et al.ACM MM 2024 · 13 citations
- PAGED: A Benchmark for Procedural Graphs Extraction from DocumentsWeihong Du, Wenrui Liao, Hongru Liang, Wenqiang LeiACL 2024 · 4 citations
