End-to-End Argumentation Knowledge Graph Construction
Khalid Al Khatib, Yufang Hou, Henning Wachsmuth, Charles Jochim, Francesca Bonin, Benno Stein
摘要
This paper studies the end-to-end construction of an NLP Knowledge Graph (KG) from scientific papers. We focus on extracting four types of relations: evaluatedOn between tasks and datasets, evaluatedBy between tasks and evaluation metrics, as well as coreferent and related relations between the same type of entities. For instance, "F1 score" is coreferent with "F-measure". We introduce novel methods for each of these relation types and apply our final framework (SciNLP-KG) to 30,000 NLP papers from ACL Anthology to build a large-scale KG, which can facilitate automatically constructing scientific leaderboards for the NLP community. The results of our experiments indicate that the resulting KG contains high-quality information. * Work done during internship at IBM Research. 1 https://github.com/sebastianruder/NLP-progress 2 https://paperswithcode.com entities. For instance, "semantic role labeling" is related to "argument identification" and "GENIA Corpus" is related to "NCBI Corpus". To evaluate our end-to-end SciNLP-KG framework, we manually construct a small-scale NLP KG based on our proposed schema (Section 4.2), which contains 85 nodes and 625 links. Experiments show that our system achieves reasonable results for all relation types on this small-scale graph with all possible meaningful links manually annotated. We further apply our framework on 30,000 NLP papers from ACL Anthology to build a largescale NLP KG containing 5,374 nodes and 15,762 relations. We evaluate the quality and coverage of the KG by manually assessing random samples and comparing it with Paperswithcode. We found that our KG contains high-quality information. Overall, the contributions of our work are threefold. First, we propose and design a new schema that represents knowledge about tasks (T), datasets (D) and metrics (M) in the NLP domain. Second, we develop a novel framework (SciNLP-KG) for constructing an NLP KG from the scientific literature in an end-to-end manner. Finally, we automatically build a large-scale NLP KG that contains high-quality information about the Task-Dataset-Metric (TDM) entities. However, our method is generalized in a way that it could be extended to the domains of computer vision or bioinformatics. Our code and datasets are made publicly available at https://github.com/Ishani-Mondal/SciKG to fuel further research.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Unsupervised stance detection for arguments from consequencesJonathan Kobbe, Ioana Hulpus, Heiner StuckenschmidtEMNLP 2020 · 被引用 27 次
- Identifying while Learning for Document Event Causality IdentificationCheng Liu, Wei Xiang, Bang WangACL 2024 · 被引用 10 次
- A Simple Contrastive Learning Framework for Interactive Argument Pair Identification via Argument-Context ExtractionLida Shi, Fausto Giunchiglia, Rui Song, Daqian Shi 等EMNLP 2022 · 被引用 4 次
- Employing Argumentation Knowledge Graphs for Neural Argument GenerationKhalid Al Khatib, Lukas Trautner, Henning Wachsmuth, Yufang Hou 等ACL 2021
- Beyond Recognising Entailment: Formalising Natural Language Inference from an Argumentative PerspectiveAmeer Saadat-Yazdi, Nadin KökciyanACL 2024
它引用的顶会 Paper1
相关 Paper
- SciNLP: A Domain-Specific Benchmark for Full-Text Scientific Entity and Relation Extraction in NLPDecheng Duan, Jitong Peng, Yingyi Zhang, Chengzhi ZhangEMNLP 2025 · 被引用 1 次
- Efficient Performance Tracking: Leveraging Large Language Models for Automated Construction of Scientific LeaderboardsFurkan Sahinuç, Thy Thy Tran, Yulia Grishina, Yufang Hou 等EMNLP 2024 · 被引用 2 次
- SciNLI: A Corpus for Natural Language Inference on Scientific TextMobashir Sadat, Cornelia CarageaACL 2022 · 被引用 41 次
- SCOPE and SCION: A Benchmark and an Auditable Reference Pipeline for Schema Induction and Fusion from TextMiaobo Hu, Xiaobo Guo, Shuhao Hu, BoKun Wang 等ICML 2026
- Can LLMs be Good Graph Judge for Knowledge Graph Construction?Haoyu Huang, Chong Chen, Zeang Sheng, Yang Li 等EMNLP 2025 · 被引用 7 次
