Learning to Extract Structured Entities Using Language Models
Haolun Wu, Ye Yuan, Liana Mikaelyan, Alexander Meulemans, Xue Liu, James Hensman, Bhaskar Mitra
摘要
Recent advances in machine learning have significantly impacted the field of information extraction, with Language Models (LMs) playing a pivotal role in extracting structured information from unstructured text. Prior works typically represent information extraction as triplet-centric and use classical metrics such as precision and recall for evaluation. We reformulate the task to be entity-centric, enabling the use of diverse metrics that can provide more insights from various perspectives. We contribute to the field by introducing Structured Entity Extraction and proposing the Approximate Entity Set OverlaP (AESOP) metric, designed to appropriately assess model performance. Later, we introduce a new Multistage Structured Entity Extraction (MuSEE) model that harnesses the power of LMs for enhanced effectiveness and efficiency by decomposing the extraction task into multiple stages. Quantitative and human side-by-side evaluations confirm that our model outperforms baselines, offering promising directions for future advancements in structured entity extraction. Our source code is available at https://github.com/microsoft/Structured-Entity-Extraction .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper4
- Efficient Streaming Language Models with Attention SinksGuangxuan Xiao, Yuandong Tian, Beidi Chen, Song Han 等ICLR 2024 · 被引用 1,714 次
- Generative Knowledge Graph Construction: A ReviewHongbin Ye, Ningyu Zhang, Hui Chen, Huajun ChenEMNLP 2022 · 被引用 51 次
- DetIE: Multilingual Open Information Extraction Inspired by Object DetectionMichael Vasilkovsky, Anton Alekseev, Valentin Malykh, Ilya Shenbin 等AAAI 2022 · 被引用 24 次
- Alignment-Augmented Consistent Translation for Multilingual Open Information ExtractionKeshav Kolluru, Muqeeth Mohammed, Shubham Mittal, Soumen Chakrabarti 等ACL 2022
相关 Paper
- ADELIE: Aligning Large Language Models on Information ExtractionYunjia Qi, Hao Peng, Xiaozhi Wang, Bin Xu 等EMNLP 2024 · 被引用 8 次
- Systematic Comparison of Neural Architectures and Training Approaches for Open Information ExtractionPatrick Hohenecker, Frank Mtumbuka, Vid Kocijan, Thomas LukasiewiczEMNLP 2020 · 被引用 10 次
- Towards Robust Information Extraction via Binomial Distribution Guided Counterpart SequenceYinhao Bai, Yuhua Zhao, Zhixin Han, Hang Gao 等KDD 2024 · 被引用 1 次
- Bridging the Gap: Aligning Language Model Generation with Structured Information Extraction via Controllable State TransitionHao Li, Yubing Ren, Yanan Cao, Yingjie Li 等WWW 2025
- Query-based Instance Discrimination Network for Relational Triple ExtractionZeqi Tan, Yongliang Shen, Xuming Hu, Wenqi Zhang 等EMNLP 2022 · 被引用 10 次
