Alternative Input Signals Ease Transfer in Multilingual Machine Translation
Simeng Sun, Angela Fan, James Cross, Vishrav Chaudhary, Chau Tran, Philipp Koehn, Francisco Guzmán
摘要
Recent work in multilingual machine translation (MMT) has focused on the potential of positive transfer between languages, particularly cases where higher-resourced languages can benefit lower-resourced ones. While training an MMT model, the supervision signals learned from one language pair can be transferred to the other via the tokens shared by multiple source languages. However, the transfer is inhibited when the token overlap among source languages is small, which manifests naturally when languages use different writing systems. In this paper, we tackle inhibited transfer by augmenting the training data with alternative signals that unify different writing systems, such as phonetic, romanized, and transliterated input. We test these signals on Indic and Turkic languages, two language families where the writing systems differ but languages still share common features. Our results indicate that a straightforward multi-source self-ensemble – training a model on a mixture of various signals and ensembling the outputs of the same model fed with different signals during inference, outperforms strong ensemble baselines by 1.3 BLEU points on both language families. Further, we find that incorporating alternative inputs via self-ensemble can be particularly effective when training set is small, leading to +5 BLEU when only 5% of the total training data is accessible. Finally, our analysis demonstrates that including alternative signals yields more consistency and translates named entities more accurately, which is crucial for increased factuality of automated systems.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Beyond Shared Vocabulary: Increasing Representational Word Similarities across Languages for Multilingual Machine TranslationDi Wu, Christof MonzEMNLP 2023 · 被引用 5 次
- Towards a Better Understanding of Variations in Zero-Shot Neural Machine Translation PerformanceShaomu Tan, Christof MonzEMNLP 2023 · 被引用 2 次
- MC²: Towards Transparent and Culturally-Aware NLP for Minority Languages in ChinaChen Zhang, Mingxu Tao, Quzhe Huang, Jiuheng Lin 等ACL 2024
- Multilingual Pixel Representations for Translation and Effective Cross-lingual TransferElizabeth Salesky, Neha Verma, Philipp Koehn, Matt PostEMNLP 2023
它引用的顶会 Paper2
相关 Paper
- HintedBT: Augmenting Back-Translation with Quality and Transliteration HintsSahana Ramnath, Melvin Johnson, Abhirut Gupta, Aravindan RaghuveerEMNLP 2021 · 被引用 7 次
- Subword Evenness (SuE) as a Predictor of Cross-lingual Transfer to Low-resource LanguagesOlga Pelloni, Anastassia Shaitarova, Tanja SamardzicEMNLP 2022 · 被引用 4 次
- Role of Language Relatedness in Multilingual Fine-tuning of Language Models: A Case Study in Indo-Aryan LanguagesTejas I. Dhamecha, V. Rudra Murthy, Samarth Bharadwaj, Karthik Sankaranarayanan 等EMNLP 2021 · 被引用 18 次
- Exploiting Language Relatedness for Low Web-Resource Language Model Adaptation: An Indic Languages StudyYash Khemchandani, Sarvesh Mehtani, Vaidehi Patil, Abhijeet Awasthi 等ACL 2021
- Knowledge Distillation for Multilingual Unsupervised Neural Machine TranslationHaipeng Sun, Rui Wang, Kehai Chen, Masao Utiyama 等ACL 2020 · 被引用 37 次
