Understanding the Properties of Minimum Bayes Risk Decoding in Neural Machine Translation
Mathias Müller, Rico Sennrich
摘要
Neural Machine Translation (NMT) currently exhibits biases such as producing translations that are too short and overgenerating frequent words, and shows poor robustness to copy noise in training data or domain shift. Recent work has tied these shortcomings to beam search -the de facto standard inference algorithm in NMT -and Eikema and Aziz (2020) propose to use Minimum Bayes Risk (MBR) decoding on unbiased samples instead. In this paper, we empirically investigate the properties of MBR decoding on a number of previously reported biases and failure cases of beam search. We find that MBR still exhibits a length and token frequency bias, owing to the MT metrics used as utility functions, but that MBR also increases robustness against copy noise in the training data and domain shift. 1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper18
- Detecting and Mitigating Hallucinations in Machine Translation: Model Internal Workings Alone Do Well, Sentence Similarity Even BetterDavid Dale, Elena Voita, Loïc Barrault, Marta R. Costa-jussàACL 2023 · 被引用 25 次
- Efficient Minimum Bayes Risk Decoding using Low-Rank Matrix Completion AlgorithmsFiras Trabelsi, David Vilar, Mara Finkelstein, Markus FreitagNeurIPS 2024 · 被引用 18 次
- Uncertainty Determines the Adequacy of the Mode and the Tractability of Decoding in Sequence-to-Sequence ModelsFelix Stahlberg, Ilia Kulikov, Shankar KumarACL 2022 · 被引用 13 次
- Sampling-Based Approximations to Minimum Bayes Risk Decoding for Neural Machine TranslationBryan Eikema, Wilker AzizEMNLP 2022 · 被引用 10 次
- Model-Based Minimum Bayes Risk Decoding for Text GenerationYuu Jinnai, Tetsuro Morimura, Ukyo Honda, Kaito Ariu 等ICML 2024 · 被引用 9 次
它引用的顶会 Paper3
- ParaCrawl: Web-Scale Acquisition of Parallel CorporaMarta Bañón, Pinzhen Chen, Barry Haddow, Kenneth Heafield 等ACL 2020 · 被引用 132 次
- SSMBA: Self-Supervised Manifold Based Data Augmentation for Improving Out-of-Domain RobustnessNathan Ng, Kyunghyun Cho, Marzyeh GhassemiEMNLP 2020 · 被引用 6 次
- COMET: A Neural Framework for MT EvaluationRicardo Rei, Craig Stewart, Ana C. Farinha, Alon LavieEMNLP 2020 · 被引用 6 次
相关 Paper
- Unveiling the Power of Source: Source-based Minimum Bayes Risk Decoding for Neural Machine TranslationBoxuan Lyu, Hidetaka Kamigaito, Kotaro Funakoshi, Manabu OkumuraACL 2025
- Quality-Aware Translation Models: Efficient Generation and Quality Estimation in a Single ModelChristian Tomani, David Vilar, Markus Freitag, Colin Cherry 等ACL 2024
- Digging Errors in NMT: Evaluating and Understanding Model Errors from Partial Hypothesis SpaceJianhao Yan, Chenming Wu, Fandong Meng, Jie ZhouEMNLP 2022
- Multi-Sentence Resampling: A Simple Approach to Alleviate Dataset Length Bias and Beam-Search DegradationIvan Provilkov, Andrey MalininEMNLP 2021 · 被引用 3 次
- MBR and QE Finetuning: Training-time Distillation of the Best and Most Expensive Decoding MethodsMara Finkelstein, Markus FreitagICLR 2024 · 被引用 39 次
