BenchIE: A Framework for Multi-Faceted Fact-Based Open Information Extraction Evaluation
Kiril Gashteovski, Mingying Yu, Bhushan Kotnis, Carolin Lawrence, Mathias Niepert, Goran Glavas
摘要
Intrinsic evaluations of OIE systems are carried out either manually—with human evaluators judging the correctness of extractions—or automatically, on standardized benchmarks. The latter, while much more cost-effective, is less reliable, primarily because of the incompleteness of the existing OIE benchmarks: the ground truth extractions do not include all acceptable variants of the same fact, leading to unreliable assessment of the models’ performance. Moreover, the existing OIE benchmarks are available for English only. In this work, we introduce BenchIE: a benchmark and evaluation framework for comprehensive evaluation of OIE systems for English, Chinese, and German. In contrast to existing OIE benchmarks, BenchIE is fact-based, i.e., it takes into account informational equivalence of extractions: our gold standard consists of fact synsets, clusters in which we exhaustively list all acceptable surface forms of the same fact. Moreover, having in mind common downstream applications for OIE, we make BenchIE multi-faceted; i.e., we create benchmark variants that focus on different facets of OIE evaluation, e.g., compactness or minimality of extractions. We benchmark several state-of-the-art OIE systems using BenchIE and demonstrate that these systems are significantly less effective than indicated by existing OIE benchmarks. We make BenchIE (data and evaluation code) publicly available.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Open Information Extraction via ChunksKuicai Dong, Aixin Sun, Jung-Jae Kim, Xiaoli LiEMNLP 2023 · 被引用 4 次
- Linking Surface Facts to Large-Scale Knowledge GraphsGorjan Radevski, Kiril Gashteovski, Chia-Chien Hung, Carolin Lawrence 等EMNLP 2023 · 被引用 2 次
- Preserving Knowledge Invariance: Rethinking Robustness Evaluation of Open Information ExtractionJi Qi, Chuchun Zhang, Xiaozhi Wang, Kaisheng Zeng 等EMNLP 2023 · 被引用 2 次
它引用的顶会 Paper8
- From Zero to Hero: On the Limitations of Zero-Shot Language Transfer with Multilingual TransformersAnne Lauscher, Vinit Ravishankar, Ivan Vulic, Goran GlavasEMNLP 2020 · 被引用 235 次
- Span Model for Open Information Extraction on Accurate CorpusJunlang Zhan, Hai ZhaoAAAI 2020 · 被引用 90 次
- KBPearl: A Knowledge Base Population System Supported by Joint Entity and Relation LinkingXueling Lin, Haoyang Li, Hao Xin, Zijian Li 等VLDB 2020 · 被引用 30 次
- Probing Linguistic Features of Sentence-Level Representations in Relation ExtractionChristoph Alt, Aleksandra Gabryszak, Leonhard HennigACL 2020 · 被引用 29 次
- Can We Predict New Facts with Open Knowledge Graph Embeddings? A Benchmark for Open Link PredictionSamuel Broscheit, Kiril Gashteovski, Yanjie Wang, Rainer GemullaACL 2020 · 被引用 27 次
相关 Paper
- When to Use What: An In-Depth Comparative Empirical Analysis of OpenIE Systems for Downstream ApplicationsKevin Pei, Ishan Jindal, Kevin Chen-Chuan Chang, ChengXiang Zhai 等ACL 2023 · 被引用 2 次
- Semi-Open Information ExtractionBowen Yu, Zhenyu Zhang, Jiawei Sheng, Tingwen Liu 等WWW 2021 · 被引用 29 次
- Systematic Comparison of Neural Architectures and Training Approaches for Open Information ExtractionPatrick Hohenecker, Frank Mtumbuka, Vid Kocijan, Thomas LukasiewiczEMNLP 2020 · 被引用 10 次
- IELM: An Open Information Extraction Benchmark for Pre-Trained Language ModelsChenguang Wang, Xiao Liu, Dawn SongEMNLP 2022 · 被引用 3 次
- DetIE: Multilingual Open Information Extraction Inspired by Object DetectionMichael Vasilkovsky, Anton Alekseev, Valentin Malykh, Ilya Shenbin 等AAAI 2022 · 被引用 24 次
