Fine-grained Information Extraction from Biomedical Literature based on Knowledge-enriched Abstract Meaning Representation
Zixuan Zhang, Nikolaus Nova Parulian, Heng Ji, Ahmed Elsayed, Skatje Myers, Martha Palmer
Abstract
Biomedical Information Extraction from scientific literature presents two unique and nontrivial challenges. First, compared with general natural language texts, sentences from scientific papers usually possess wider contexts between knowledge elements. Moreover, comprehending the fine-grained scientific entities and events urgently requires domain-specific background knowledge. In this paper, we propose a novel biomedical Information Extraction (IE) model to tackle these two challenges and extract scientific entities and events from English research papers. We perform Abstract Meaning Representation (AMR) to compress the wide context to uncover a clear semantic structure for each complex sentence. Besides, we construct the sentence-level knowledge graph from an external knowledge base and use it to enrich the AMR graph to improve the model's understanding of complex scientific concepts. We use an edge-conditioned graph attention network to encode the knowledgeenriched AMR graph for biomedical IE tasks. Experiments on the GENIA 2011 dataset show that the AMR and external knowledge have contributed 1.8% and 3.0% absolute F-score gains respectively. In order to evaluate the impact of our approach on real-world problems that involve topic-specific fine-grained knowledge elements, we have also created a new ontology and annotated corpus for entity and event extraction for the COVID-19 scientific literature, which can serve as a new benchmark for the biomedical IE community. 1
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext fa569b66-6b67-41d0-a095-6d9f3d8e7893Cited by top-tier papers3
- Text2Mol: Cross-Modal Molecule Retrieval with Natural Language QueriesCarl Edwards, ChengXiang Zhai, Heng JiEMNLP 2021 · 79 citations
- A Survey of AMR ApplicationsShira Wein, Juri OpitzEMNLP 2024 · 7 citations
- Linguistic representations for fewer-shot relation extraction across domainsSireesh Gururaja, Ritam Dutt, Tinglong Liao, Carolyn P. RoséACL 2023 · 5 citations
Builds on1
Related papers
- Joint Biomedical Entity and Relation Extraction with Knowledge-Enhanced Collective InferenceTuan Manh Lai, Heng Ji, ChengXiang Zhai, Quan Hung TranACL 2021
- Improving Biomedical Abstractive Summarisation with Knowledge Aggregation from Citation PapersChen Tang, Shun Wang, Tomas Goldsack, Chenghua LinEMNLP 2023 · 5 citations
- Bio-RFX: Refining Biomedical Extraction via Advanced Relation Classification and Structural ConstraintsMinjia Wang, Fangzhou Liu, Xiuxing Li, Bowen Dong et al.EMNLP 2024 · 1 citation
- Learning Conceptual-Contextual Embeddings for Medical TextXiao Zhang, Dejing Dou, Ji WuAAAI 2020 · 17 citations
- Knowledge-Graph Augmented Word Representations for Named Entity RecognitionQizhen He, Liang Wu, Yida Yin, Heming CaiAAAI 2020 · 30 citations
