Generating Diverse Translation from Model Distribution with Dropout
Xuanfu Wu, Yang Feng, Chenze Shao
Abstract
Despite the improvement of translation quality, neural machine translation (NMT) often suffers from the lack of diversity in its generation. In this paper, we propose to generate diverse translations by deriving a large number of possible models with Bayesian modelling and sampling models from them for inference. The possible models are obtained by applying concrete dropout to the NMT model and each of them has specific confidence for its prediction, which corresponds to a posterior model distribution under specific training data in the principle of Bayesian modeling. With variational inference, the posterior model distribution can be approximated with a variational distribution, from which the final models for inference are sampled. We conducted experiments on Chinese-English and English-German translation tasks and the results shows that our method makes a better trade-off between diversity and accuracy.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ea1a0a46-8068-49ab-86e5-5fcc0c9f5b5bCited by top-tier papers2
- WeTS: A Benchmark for Translation SuggestionZhen Yang, Fandong Meng, Yingxue Zhang, Ernan Li et al.EMNLP 2022 · 5 citations
- EAG: Extract and Generate Multi-way Aligned Corpus for Complete Multi-lingual Neural Machine TranslationYulin Xu, Zhen Yang, Fandong Meng, Jie ZhouACL 2022 · 3 citations
Builds on3
- Reducing Transformer Depth on Demand with Structured DropoutAngela Fan, Edouard Grave, Armand JoulinICLR 2020 · 695 citations
- Generating Diverse Translation by Manipulating Multi-Head AttentionZewei Sun, Shujian Huang, Hao-Ran Wei, Xinyu Dai et al.AAAI 2020 · 36 citations
- Modeling Fluency and Faithfulness for Diverse Neural Machine TranslationYang Feng, Wanying Xie, Shuhao Gu, Chenze Shao et al.AAAI 2020 · 28 citations
Related papers
- Bayesian Posterior Approximation With Stochastic EnsemblesOleksandr Balabanov, Bernhard Mehlig, Hampus LinanderCVPR 2023
- DIBS: Diversity Inducing Information Bottleneck in Model EnsemblesSamarth Sinha, Homanga Bharadhwaj, Anirudh Goyal, Hugo Larochelle et al.AAAI 2021 · 43 citations
- BARNN: A Bayesian Autoregressive and Recurrent Neural NetworkDario Coscia, Max Welling, Nicola Demo, Gianluigi RozzaICML 2025
- Data Diversification: A Simple Strategy For Neural Machine TranslationXuan-Phi Nguyen, Shafiq R. Joty, Kui Wu, Ai Ti AwNeurIPS 2020 · 75 citations
- Contextual Dropout: An Efficient Sample-Dependent Dropout ModuleXinjie Fan, Shujian Zhang, Korawat Tanwisuth, Xiaoning Qian et al.ICLR 2021 · 34 citations
