CitationIE: Leveraging the Citation Graph for Scientific Information Extraction
Vijay Viswanathan, Graham Neubig, Pengfei Liu
摘要
Automatically extracting key information from scientific documents has the potential to help scientists work more efficiently and accelerate the pace of scientific progress. Prior work has considered extracting documentlevel entity clusters and relations end-to-end from raw scientific text, which can improve literature search and help identify methods and materials for a given problem. Despite the importance of this task, most existing works on scientific information extraction (SciIE) consider extraction solely based on the content of an individual paper, without considering the paper's place in the broader literature. In contrast to prior work, we augment our text representations by leveraging a complementary source of document context: the citation graph of referential links between citing and cited papers. On a test set of English-language scientific documents, we show that simple ways of utilizing the structure and content of the citation graph can each lead to significant gains in different scientific information extraction tasks. When these tasks are combined, we observe a sizable improvement in end-to-end information extraction over the state-of-the-art, suggesting the potential for future work along this direction. We release software tools to facilitate citation-aware SciIE development. 1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Text2NKG: Fine-Grained N-ary Relation Extraction for N-ary relational Knowledge Graph ConstructionHaoran Luo, Haihong E, Yuhao Yang, Tianyu Yao 等NeurIPS 2024 · 被引用 19 次
- ReSel: N-ary Relation Extraction from Scientific Text and Tables by Learning to Retrieve and SelectYuchen Zhuang, Yinghao Li, Junyang Zhang, Yue Yu 等EMNLP 2022 · 被引用 11 次
- SciER: An Entity and Relation Extraction Dataset for Datasets, Methods, and Tasks in Scientific DocumentsQi Zhang, Zhijia Chen, Huitong Pan, Cornelia Caragea 等EMNLP 2024 · 被引用 7 次
- GSAP-ERE: Fine-Grained Scholarly Entity and Relation Extraction Focused on Machine LearningWolfgang Otto, Lu Gan, Sharmila Upadhyaya, Saurav Karmakar 等AAAI 2026
- Content- and Topology-Aware Representation Learning for Scientific Multi-LiteratureKai Zhang, Kaisong Song, Yangyang Kang, Xiaozhong LiuEMNLP 2023
它引用的顶会 Paper3
- S2ORC: The Semantic Scholar Open Research CorpusKyle Lo, Lucy Lu Wang, Mark Neumann, Rodney Kinney 等ACL 2020 · 被引用 424 次
- SciREX: A Challenge Dataset for Document-Level Information ExtractionSarthak Jain, Madeleine van Zuylen, Hannaneh Hajishirzi, Iz BeltagyACL 2020 · 被引用 9 次
- AxCell: Automatic Extraction of Results from Machine Learning PapersMarcin Kardas, Piotr Czapla, Pontus Stenetorp, Sebastian Ruder 等EMNLP 2020 · 被引用 5 次
相关 Paper
- DisenCite: Graph-Based Disentangled Representation Learning for Context-Specific Citation GenerationYifan Wang, Yiping Song, Shuai Li, Chaoran Cheng 等AAAI 2022 · 被引用 42 次
- Scientific Paper Extractive Summarization Enhanced by Citation GraphsXiuying Chen, Mingzhe Li, Shen Gao, Rui Yan 等EMNLP 2022 · 被引用 8 次
- CitationSum: Citation-aware Graph Contrastive Learning for Scientific Paper SummarizationZheheng Luo, Qianqian Xie, Sophia AnaniadouWWW 2023 · 被引用 19 次
- Fine-grained Information Extraction from Biomedical Literature based on Knowledge-enriched Abstract Meaning RepresentationZixuan Zhang, Nikolaus Nova Parulian, Heng Ji, Ahmed Elsayed 等ACL 2021
- SPECTER: Document-level Representation Learning using Citation-informed TransformersArman Cohan, Sergey Feldman, Iz Beltagy, Doug Downey 等ACL 2020 · 被引用 20 次
