MetaMT, a Meta Learning Method Leveraging Multiple Domain Data for Low Resource Machine Translation
Rumeng Li, Xun Wang, Hong Yu
Abstract
Neural machine translation (NMT) models have achieved state-of-the-art translation quality with a large quantity of parallel corpora available. However, their performance suffers significantly when it comes to domain-specific translations, in which training data are usually scarce. In this paper, we present a novel NMT model with a new word embedding transition technique for fast domain adaption. We propose to split parameters in the model into two groups: model parameters and meta parameters. The former are used to model the translation while the latter are used to adjust the representational space to generalize the model to different domains. We mimic the domain adaptation of the machine translation model to low-resource domains using multiple translation tasks on different domains. A new training strategy based on meta-learning is developed along with the proposed model to update the model parameters and meta parameters alternately. Experiments on datasets of different domains showed substantial improvements of NMT performances on a limited amount of data.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a57c16e7-22b8-49f5-8c2a-17287a605a87Cited by top-tier papers9
- Meta-Curriculum Learning for Domain Adaptation in Neural Machine TranslationRunzhe Zhan, Xuebo Liu, Derek F. Wong, Lidia S. ChaoAAAI 2021 · 50 citations
- Meta-Transfer Learning for Low-Resource Abstractive SummarizationYi-Syuan Chen, Hong-Han ShuaiAAAI 2021 · 41 citations
- Improving Meta-learning for Low-resource Text Classification and Generation via Memory ImitationYingxiu Zhao, Zhiliang Tian, Huaxiu Yao, Yinhe Zheng et al.ACL 2022 · 21 citations
- GLUECons: A Generic Benchmark for Learning under ConstraintsHossein Rajaby Faghihi, Aliakbar Nafar, Chen Zheng, Roshanak Mirzaee et al.AAAI 2023 · 18 citations
- Learning a Gradient-free Riemannian Optimizer on Tangent SpacesXiaomeng Fan, Zhi Gao, Yuwei Wu, Yunde Jia et al.AAAI 2021 · 8 citations
Related papers
- Unsupervised Neural Machine Translation for Low-Resource Domains via Meta-LearningCheonbok Park, Yunwon Tae, Taehee Kim, Soyoung Yang et al.ACL 2021
- Distilling Multiple Domains for Neural Machine TranslationAnna Currey, Prashant Mathur, Georgiana DinuEMNLP 2020 · 19 citations
- Meta Back-TranslationHieu Pham, Xinyi Wang, Yiming Yang, Graham NeubigICLR 2021 · 26 citations
- MetaNER: Named Entity Recognition with Meta-LearningJing Li, Shuo Shang, Ling ShaoWWW 2020 · 56 citations
- Domain adapted machine translation: What does catastrophic forgetting forget and why?Danielle Saunders, Steve DeNeefeEMNLP 2024 · 1 citation
