Continual Knowledge Distillation for Neural Machine Translation
Yuanchi Zhang, Peng Li, Maosong Sun, Yang Liu
Abstract
While many parallel corpora are not publicly accessible for data copyright, data privacy and competitive differentiation reasons, trained translation models are increasingly available on open platforms. In this work, we propose a method called continual knowledge distillation to take advantage of existing translation models to improve one model of interest. The basic idea is to sequentially transfer knowledge from each trained model to the distilled model. Extensive experiments on Chinese-English and German-English datasets show that our method achieves significant and consistent improvements over strong baselines under both homogeneous and heterogeneous trained model settings and is robust to malicious models. 1
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers1
Ask how each one uses itBuilds on5
- Improved Knowledge Distillation via Teacher AssistantSeyed-Iman Mirzadeh, Mehrdad Farajtabar, Ang Li, Nir Levine et al.AAAI 2020 · 1,361 citations
- Finding Sparse Structures for Domain Specific Neural Machine TranslationJianze Liang, Chengqi Zhao, Mingxuan Wang, Xipeng Qiu et al.AAAI 2021 · 33 citations
- Lifelong Language Knowledge DistillationYung-Sung Chuang, Shang-Yu Su, Yun-Nung ChenEMNLP 2020 · 33 citations
- Entropy-Based Vocabulary Substitution for Incremental Learning in Multilingual Neural Machine TranslationKaiyu Huang, Peng Li, Jin Ma, Yang LiuEMNLP 2022 · 7 citations
- Selective Knowledge Distillation for Neural Machine TranslationFusheng Wang, Jianhao Yan, Fandong Meng, Jie ZhouACL 2021
Related papers
- Knowledge Transfer in Incremental Learning for Multilingual Neural Machine TranslationKaiyu Huang, Peng Li, Jin Ma, Ting Yao et al.ACL 2023 · 17 citations
- Continual Learning with Semi-supervised Contrastive Distillation for Incremental Neural Machine TranslationYunlong Liang, Fandong Meng, Jiaan Wang, Jinan Xu et al.ACL 2024 · 7 citations
- Overcoming Catastrophic Forgetting beyond Continual Learning: Balanced Training for Neural Machine TranslationChenze Shao, Yang FengACL 2022 · 40 citations
- Pretrained Bidirectional Distillation for Machine TranslationYimeng Zhuang, Mei TuACL 2023 · 3 citations
- Acquiring Knowledge from Pre-Trained Model to Neural Machine TranslationRongxiang Weng, Heng Yu, Shujian Huang, Shanbo Cheng et al.AAAI 2020 · 71 citations
