Partial Colexifications Improve Concept Embeddings
Arne Rubehn, Johann-Mattis List
摘要
While the embedding of words has revolutionized the field of Natural Language Processing, the embedding of concepts has received much less attention so far. A dense and meaningful representation of concepts, however, could prove useful for several tasks in computational linguistics, especially those involving cross-linguistic data or sparse data from low resource languages. First methods that have been proposed so far embed concepts from automatically constructed colexification networks. While these approaches depart from automatically inferred polysemies, attested across a larger number of languages, they are restricted to the word level, ignoring lexical relations that would only hold for parts of the words in a given language. Building on recently introduced methods for the inference of partial colexifications, we show how they can be used to improve concept embeddings in meaningful ways. The learned embeddings are evaluated against lexical similarity ratings, recorded instances of semantic shift, and word association data. We show that in all evaluation tasks, the inclusion of partial colexifications lead to improved concept representations and better results. Our results further show that the learned embeddings are able to capture and represent different semantic relationships between concepts.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper1
相关 Paper
- To Word Senses and Beyond: Inducing Concepts with Contextualized Language ModelsBastien Liétard, Pascal Denis, Mikaela KellerEMNLP 2024
- Generalising Multilingual Concept-to-Text NLG with Language Agnostic DelexicalisationGiulio Zhou, Gerasimos LampourasACL 2021
- A Comparison of Architectures and Pretraining Methods for Contextualized Multilingual Word EmbeddingsNiels van der Heijden, Samira Abnar, Ekaterina ShutovaAAAI 2020 · 被引用 16 次
- Unsupervised Interlingual Semantic Representations from Sentence Embeddings for Zero-Shot Cross-Lingual TransferChanny Hong, Jaeyeon Lee, Jungkwon LeeAAAI 2020 · 被引用 1 次
- Distilling Semantic Concept Embeddings from Contrastively Fine-Tuned Language ModelsNa Li, Hanane Kteich, Zied Bouraoui, Steven SchockaertSIGIR 2023 · 被引用 3 次
