Interpretable Word Sense Representations via Definition Generation: The Case of Semantic Change Analysis
Mario Giulianelli, Iris Luden, Raquel Fernández, Andrey Kutuzov
摘要
We propose using automatically generated natural language definitions of contextualised word usages as interpretable word and word sense representations.Given a collection of usage examples for a target word, and the corresponding data-driven usage clusters (i.e., word senses), a definition is generated for each usage with a specialised Flan-T5 language model, and the most prototypical definition in a usage cluster is chosen as the sense label. We demonstrate how the resulting sense labels can make existing approaches to semantic change analysis more interpretable, and how they can allow users — historical linguists, lexicographers, or social scientists — to explore and intuitively explain diachronic trajectories of word meaning. Semantic change analysis is only one of many possible applications of the ‘definitions as representations’ paradigm. Beyond being human-readable, contextualised definitions also outperform token or usage sentence embeddings in word-in-context semantic similarity judgements, making them a new promising type of lexical representation for NLP.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Using Synchronic Definitions and Semantic Relations to Classify Semantic Change TypesPierluigi Cassotti, Stefano De Pascale, Nina TahmasebiACL 2024 · 被引用 2 次
- More DWUGs: Extending and Evaluating Word Usage Graph Datasets in Multiple LanguagesDominik Schlechtweg, Pierluigi Cassotti, Bill Noble, David Alfter 等EMNLP 2024 · 被引用 2 次
- Hateful Word in Context ClassificationSanne Hoeken, Sina Zarrieß, Özge AlaçamEMNLP 2024 · 被引用 2 次
- Automatically Generated Definitions and their utility for Modeling Word MeaningFrancesco Periti, David Alfter, Nina TahmasebiEMNLP 2024 · 被引用 2 次
- WSDPO: A Generative Word Sense Disambiguation Framework with Chain-of-Thought and Preference OptimizationKunpeng Kang, Shuaimin Li, Kaiyuan Zhang, Luyang Zhang 等ACL 2026
它引用的顶会 Paper7
- Generationary or "How We Went beyond Word Sense Inventories and Learned to Gloss"Michele Bevilacqua, Marco Maru, Roberto NavigliEMNLP 2020 · 被引用 41 次
- Understanding Jargon: Combining Extraction and Generation for Definition ModelingJie Huang, Hanyin Shao, Kevin Chen-Chuan Chang, Jinjun Xiong 等EMNLP 2022 · 被引用 11 次
- DWUG: A large Resource of Diachronic Word Usage Graphs in Four LanguagesDominik Schlechtweg, Nina Tahmasebi, Simon Hengchen, Haim Dubossarsky 等EMNLP 2021 · 被引用 1 次
- Multitasking Framework for Unsupervised Simple Definition GenerationCunliang Kong, Yun Chen, Hengyuan Zhang, Liner Yang 等ACL 2022
- Lexical Semantic Change DiscoverySinan Kurtyigit, Maike Park, Dominik Schlechtweg, Jonas Kuhn 等ACL 2021
相关 Paper
- Analysing Lexical Semantic Change with Contextualised Word RepresentationsMario Giulianelli, Marco Del Tredici, Raquel FernándezACL 2020 · 被引用 118 次
- Analyzing Semantic Change through Lexical ReplacementsFrancesco Periti, Pierluigi Cassotti, Haim Dubossarsky, Nina TahmasebiACL 2024
- Detecting Contact-Induced Semantic Shifts: What Can Embedding-Based Methods Do in Practice?Filip Miletic, Anne Przewozny-Desriaux, Ludovic TanguyEMNLP 2021 · 被引用 5 次
- SensEmBERT: Context-Enhanced Sense Embeddings for Multilingual Word Sense DisambiguationBianca Scarlini, Tommaso Pasini, Roberto NavigliAAAI 2020 · 被引用 121 次
- Word2Fun: Modelling Words as Functions for Diachronic Word RepresentationBenyou Wang, Emanuele Di Buccio, Massimo MelucciNeurIPS 2021 · 被引用 6 次
