Beyond Shared Vocabulary: Increasing Representational Word Similarities across Languages for Multilingual Machine Translation
Di Wu, Christof Monz
摘要
Using a shared vocabulary is common practice in Multilingual Neural Machine Translation (MNMT). In addition to its simple design, shared tokens play an important role in positive knowledge transfer, which manifests naturally when the shared tokens refer to similar meanings across languages. However, when words overlap is small, e.g., using different writing systems, transfer is inhibited. In this paper, we propose a re-parameterized method for building embeddings to alleviate this problem. More specifically, we define word-level information transfer pathways via word equivalence classes and rely on graph networks to fuse word embeddings across languages. Our experiments demonstrate the advantages of our approach: 1) the semantics of embeddings are better aligned across languages, 2) our method achieves evident BLEU improvements on high- and low-resource MNMT, and 3) only less than 1.0% additional trainable parameters are required with a limited increase in computational costs, while the inference time is identical to baselines.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Towards a Better Understanding of Variations in Zero-Shot Neural Machine Translation PerformanceShaomu Tan, Christof MonzEMNLP 2023 · 被引用 2 次
- Neuron Specialization: Leveraging Intrinsic Task Modularity for Multilingual Machine TranslationShaomu Tan, Di Wu, Christof MonzEMNLP 2024
它引用的顶会 Paper5
- In Neural Machine Translation, What Does Transfer Learning Transfer?Alham Fikri Aji, Nikolay Bogoychev, Kenneth Heafield, Rico SennrichACL 2020 · 被引用 56 次
- Alternative Input Signals Ease Transfer in Multilingual Machine TranslationSimeng Sun, Angela Fan, James Cross, Vishrav Chaudhary 等ACL 2022 · 被引用 18 次
- COMET: A Neural Framework for MT EvaluationRicardo Rei, Craig Stewart, Ana C. Farinha, Alon LavieEMNLP 2020 · 被引用 6 次
- Towards a Better Understanding of Variations in Zero-Shot Neural Machine Translation PerformanceShaomu Tan, Christof MonzEMNLP 2023 · 被引用 2 次
- Learning Language Specific Sub-network for Multilingual Machine TranslationZehui Lin, Liwei Wu, Mingxuan Wang, Lei LiACL 2021
相关 Paper
- Multilingual Machine Translation with Hyper-AdaptersChristos Baziotis, Mikel Artetxe, James Cross, Shruti BhosaleEMNLP 2022 · 被引用 20 次
- Adaptive Token-level Cross-lingual Feature Mixing for Multilingual Neural Machine TranslationJunpeng Liu, Kaiyu Huang, Jiuyi Li, Huan Liu 等EMNLP 2022 · 被引用 5 次
- A Comparison of Architectures and Pretraining Methods for Contextualized Multilingual Word EmbeddingsNiels van der Heijden, Samira Abnar, Ekaterina ShutovaAAAI 2020 · 被引用 16 次
- Graph-Based Multilingual Label Propagation for Low-Resource Part-of-Speech TaggingAyyoob Imani, Silvia Severini, Masoud Jalili Sabet, François Yvon 等EMNLP 2022 · 被引用 8 次
- Emerging Cross-lingual Structure in Pretrained Language ModelsAlexis Conneau, Shijie Wu, Haoran Li, Luke Zettlemoyer 等ACL 2020 · 被引用 210 次
