GENEVA: Benchmarking Generalizability for Event Argument Extraction with Hundreds of Event Types and Argument Roles
Tanmay Parekh, I-Hung Hsu, Kuan-Hao Huang, Kai-Wei Chang, Nanyun Peng
摘要
Recent works in Event Argument Extraction (EAE) have focused on improving model generalizability to cater to new events and domains. However, standard benchmarking datasets like ACE and ERE cover less than 40 event types and 25 entity-centric argument roles. Limited diversity and coverage hinder these datasets from adequately evaluating the generalizability of EAE models. In this paper, we first contribute by creating a large and diverse EAE ontology. This ontology is created by transforming FrameNet, a comprehensive semantic role labeling (SRL) dataset for EAE, by exploiting the similarity between these two tasks. Then, exhaustive human expert annotations are collected to build the ontology, concluding with 115 events and 220 argument roles, with a significant portion of roles not being entities. We utilize this ontology to further introduce GENEVA, a diverse generalizability benchmarking dataset comprising four test suites, aimed at evaluating models' ability to handle limited data and unseen event type generalization. We benchmark six EAE models from various families. The results show that owing to non-entity argument roles, even the best-performing model can only achieve 39% F1 score, indicating how GENEVA provides new challenges for generalization in EAE. Overall, our large and diverse EAE ontology can aid in creating more comprehensive future resources, while GENEVA is a challenging benchmarking dataset encouraging further research for improving generalizability in EAE. The code and data can be found at https: //github.com/PlusLabNLP/GENEVA .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Improving Event Definition Following For Zero-Shot Event DetectionZefan Cai, Po-Nien Kung, Ashima Suvarna, Mingyu Derek Ma 等ACL 2024 · 被引用 3 次
- SNaRe: Domain-aware Data Generation for Low-Resource Event DetectionTanmay Parekh, Yuxuan Dong, Lucas Bandarkar, Artin Kim 等EMNLP 2025 · 被引用 1 次
- DiCoRe: Enhancing Zero-shot Event Detection via Divergent-Convergent LLM ReasoningTanmay Parekh, Kartik Mehta, Ninareh Mehrabi, Kai-Wei Chang 等EMNLP 2025 · 被引用 1 次
- Document-Level Event-Argument Data Augmentation for Challenging Role TypesJoseph Gatto, Omar Sharif, Parker Seegmiller, Sarah Masud PreumACL 2025
它引用的顶会 Paper8
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Event Extraction by Answering (Almost) Natural QuestionsXinya Du, Claire CardieEMNLP 2020 · 被引用 391 次
- A Joint Neural Model for Information Extraction with Global FeaturesYing Lin, Heng Ji, Fei Huang, Lingfei WuACL 2020 · 被引用 376 次
- Event Extraction as Machine Reading ComprehensionJian Liu, Yubo Chen, Kang Liu, Wei Bi 等EMNLP 2020 · 被引用 300 次
- ASER: A Large-scale Eventuality Knowledge GraphHongming Zhang, Xin Liu, Haojie Pan, Yangqiu Song 等WWW 2020 · 被引用 183 次
相关 Paper
- MAVEN-ARG: Completing the Puzzle of All-in-One Event Understanding Dataset with Event Argument AnnotationXiaozhi Wang, Hao Peng, Yong Guan, Kaisheng Zeng 等ACL 2024
- Few-Shot Document-Level Event Argument ExtractionXianjun Yang, Yujie Lu, Linda R. PetzoldACL 2023 · 被引用 5 次
- Query Your Model with Definitions in FrameNet: An Effective Method for Frame Semantic Role LabelingCe Zheng, Yiming Wang, Baobao ChangAAAI 2023 · 被引用 6 次
- Transfer Learning from Semantic Role Labeling to Event Argument Extraction with Template-based Slot QueryingZhisong Zhang, Emma Strubell, Eduard H. HovyEMNLP 2022 · 被引用 4 次
- GLEN: General-Purpose Event Detection for Thousands of TypesSha Li, Qiusi Zhan, Kathryn Conger, Martha Palmer 等EMNLP 2023 · 被引用 9 次
