Understanding Jargon: Combining Extraction and Generation for Definition Modeling
Jie Huang, Hanyin Shao, Kevin Chen-Chuan Chang, Jinjun Xiong, Wen-Mei Hwu
摘要
Can machines know what twin prime is? From the composition of this phrase, machines may guess twin prime is a certain kind of prime, but it is still difficult to deduce exactly what twin stands for without additional knowledge. Here, twin prime is a jargon - a specialized term used by experts in a particular field. Explaining jargon is challenging since it usually requires domain knowledge to understand. Recently, there is an increasing interest in extracting and generating definitions of words automatically. However, existing approaches, either extraction or generation, perform poorly on jargon. In this paper, we propose to combine extraction and generation for jargon definition modeling: first extract self- and correlative definitional information of target jargon from the Web and then generate the final definitions by incorporating the extracted definitional information. Our framework is remarkably simple but effective: experiments demonstrate our method can generate high-quality definitions for jargon and outperform state-of-the-art models significantly, e.g., BLEU score from 8.76 to 22.66 and human-annotated score from 2.34 to 4.04.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Explainable Notes: Examining How to Unlock Meaning in Medical Notes with Interactivity and Artificial IntelligenceHita Kambhamettu, Danaë Metaxa, Kevin Johnson, Andrew HeadCHI 2024 · 被引用 16 次
- Interpretable Word Sense Representations via Definition Generation: The Case of Semantic Change AnalysisMario Giulianelli, Iris Luden, Raquel Fernández, Andrey KutuzovACL 2023 · 被引用 9 次
- DEER: Descriptive Knowledge Graph for Explaining Entity RelationshipsJie Huang, Kerui Zhu, Kevin Chen-Chuan Chang, Jinjun Xiong 等EMNLP 2022 · 被引用 8 次
- DimonGen: Diversified Generative Commonsense Reasoning for Explaining Concept RelationshipsChenzhengyi Liu, Jie Huang, Kerui Zhu, Kevin Chen-Chuan ChangACL 2023 · 被引用 6 次
- MedReadMe: A Systematic Study for Fine-grained Sentence Readability in Medical DomainChao Jiang, Wei XuEMNLP 2024 · 被引用 3 次
它引用的顶会 Paper8
- BERTScore: Evaluating Text Generation with BERTTianyi Zhang, Varsha Kishore, Felix Wu, Kilian Q. Weinberger 等ICLR 2020 · 被引用 8,443 次
- A Joint Model for Definition Extraction with Syntactic Connection and Semantic ConsistencyAmir Pouran Ben Veyseh, Franck Dernoncourt, Dejing Dou, Thien Huu NguyenAAAI 2020 · 被引用 43 次
- Generationary or "How We Went beyond Word Sense Inventories and Learned to Gloss"Michele Bevilacqua, Marco Maru, Roberto NavigliEMNLP 2020 · 被引用 41 次
- Explicit Semantic Decomposition for Definition GenerationJiahuan Li, Yu Bao, Shujian Huang, Xinyu Dai 等ACL 2020 · 被引用 14 次
- VCDM: Leveraging Variational Bi-encoding and Deep Contextualized Word Representations for Improved Definition ModelingMachel Reid, Edison Marrese-Taylor, Yutaka MatsuoEMNLP 2020 · 被引用 14 次
相关 Paper
- MedJEx: A Medical Jargon Extraction Model with Wiki's Hyperlink Span and Contextualized Masked Language Model ScoreSunjae Kwon, Zonghai Yao, Harmon S. Jordan, David A. Levy 等EMNLP 2022 · 被引用 12 次
- Generating and Visualizing Trace Link ExplanationsYalin Liu, Jinfeng Lin, Oghenemaro Anuyah, Ronald A. Metoyer 等ICSE 2022 · 被引用 4 次
- Incorporating medical knowledge in BERT for clinical relation extractionArpita Roy, Shimei PanEMNLP 2021 · 被引用 56 次
- Probing Simile Knowledge from Pre-trained Language ModelsWeijie Chen, Yongzhu Chang, Rongsheng Zhang, Jiashu Pu 等ACL 2022
- ExPUNations: Augmenting Puns with Keywords and ExplanationsJiao Sun, Anjali Narayan-Chen, Shereen Oraby, Alessandra Cervone 等EMNLP 2022 · 被引用 7 次
