LFDe: A Lighter, Faster and More Data-Efficient Pre-training Framework for Event Extraction
Zhigang Kan, Liwen Peng, Yifu Gao, Ning Liu, Linbo Qiao, Dongsheng Li
摘要
Pre-training Event Extraction (EE) models on unlabeled data is an effective strategy that frees researchers from costly and labor-intensive data annotation. However, existing pre-training methods necessitate substantial computational resources, requiring high-performance hardware infrastructure and extensive training duration. In response to these challenges, this paper proposes a Lighter, Faster, and more Data-efficient pre-training framework for EE, named LFDe. Distinct from existing methods that strive to establish a comprehensive representation space during pre-training, our framework focuses on quickly familiarizing with the task format from a small amount of automatically constructed pseudo-events. It comprises three stages: weak-label data construction, pre-training, and fine-tuning. Specifically, during the first stage, LFDe first automatically designates pseudo-triggers and arguments based on the characteristics of real events to form pre-training samples. In the processes of pre-training and fine-tuning, the framework reframes EE as the identification of tokens semantically closest to the prompt within the given sentence. This paper also introduces a novel prompt-based sequence labeling model for EE to accommodate this reframing. Experiments on real-world datasets show that compared to similar models, our framework requires fewer pre-training data (only about 0.04%), a shorter pre-training period (about 0.03%), and lower memory requirements (about 57.6%). Simultaneously, our framework significantly improves performance in various data-scarce scenarios.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper1
问问它们各自怎么用它相关 Paper
- Prompt for Extraction? PAIE: Prompting Argument Interaction for Event Argument ExtractionYubo Ma, Zehao Wang, Yixin Cao, Mukai Li 等ACL 2022 · 被引用 182 次
- CLEVE: Contrastive Pre-training for Event ExtractionZiqi Wang, Xiaozhi Wang, Xu Han, Yankai Lin 等ACL 2021
- Unleash GPT-2 Power for Event DetectionAmir Pouran Ben Veyseh, Viet Dac Lai, Franck Dernoncourt, Thien Huu NguyenACL 2021
- MatchPrompt: Prompt-based Open Relation Extraction with Semantic Consistency Guided ClusteringJiaxin Wang, Lingling Zhang, Jun Liu, Xi Liang 等EMNLP 2022 · 被引用 4 次
- Revisiting Event Argument Extraction: Can EAE Models Learn Better When Being Aware of Event Co-occurrences?Yuxin He, Jingyue Hu, Buzhou TangACL 2023 · 被引用 26 次
