Multi-Domain Neural Machine Translation with Word-Level Adaptive Layer-wise Domain Mixing
Haoming Jiang, Chen Liang, Chong Wang, Tuo Zhao
摘要
Many multi-domain neural machine translation (NMT) models achieve knowledge transfer by enforcing one encoder to learn shared embedding across domains. However, this design lacks adaptation to individual domains. To overcome this limitation, we propose a novel multi-domain NMT model using individual modules for each domain, on which we apply word-level, adaptive and layer-wise domain mixing. We first observe that words in a sentence are often related to multiple domains. Hence, we assume each word has a domain proportion, which indicates its domain preference. Then word representations are obtained by mixing their embedding in individual domains based on their domain proportions. We show this can be achieved by carefully designing multi-head dot-product attention modules for different domains, and eventually taking weighted averages of their parameters by word-level layer-wise domain proportions. Through this, we can achieve effective domain knowledge sharing, and capture fine-grained domain-specific knowledge as well. Our experiments show that our proposed model outperforms existing ones in several NMT tasks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- CAMERO: Consistency Regularized Ensemble of Perturbed Language Models with Weight SharingChen Liang, Pengcheng He, Yelong Shen, Weizhu Chen 等ACL 2022 · 被引用 6 次
- Adaptive Token-level Cross-lingual Feature Mixing for Multilingual Neural Machine TranslationJunpeng Liu, Kaiyu Huang, Jiuyi Li, Huan Liu 等EMNLP 2022 · 被引用 5 次
- EnsLM: Ensemble Language Model for Data Diversity by Semantic ClusteringZhibin Duan, Hao Zhang, Chaojie Wang, Zhengjue Wang 等ACL 2021
- DMDTEval: An Evaluation and Analysis of LLMs on Disambiguation in Multi-domain TranslationZhibo Man, Yuanmeng Chen, Yujie Zhang, Jinan XuEMNLP 2025
它引用的顶会 Paper1
相关 Paper
- Go From the General to the Particular: Multi-Domain Translation with Domain Transformation NetworksYong Wang, Longyue Wang, Shuming Shi, Victor O. K. Li 等AAAI 2020 · 被引用 30 次
- Distilling Multiple Domains for Neural Machine TranslationAnna Currey, Prashant Mathur, Georgiana DinuEMNLP 2020 · 被引用 19 次
- MetaMT, a Meta Learning Method Leveraging Multiple Domain Data for Low Resource Machine TranslationRumeng Li, Xun Wang, Hong YuAAAI 2020 · 被引用 42 次
- Pay Better Attention to Attention: Head Selection in Multilingual and Multi-Domain Sequence ModelingHongyu Gong, Yun Tang, Juan Miguel Pino, Xian LiNeurIPS 2021 · 被引用 15 次
- Parameter Differentiation Based Multilingual Neural Machine TranslationQian Wang, Jiajun ZhangAAAI 2022 · 被引用 21 次
