Detecting Word Sense Disambiguation Biases in Machine Translation for Model-Agnostic Adversarial Attacks
Denis Emelin, Ivan Titov, Rico Sennrich
摘要
Word sense disambiguation is a well-known source of translation errors in NMT. We posit that some of the incorrect disambiguation choices are due to models' over-reliance on dataset artifacts found in training data, specifically superficial word co-occurrences, rather than a deeper understanding of the source text. We introduce a method for the prediction of disambiguation errors based on statistical data properties, demonstrating its effectiveness across several domains and model types. Moreover, we develop a simple adversarial attack strategy that minimally perturbs sentences in order to elicit disambiguation errors to further probe the robustness of translation models. Our findings indicate that disambiguation robustness varies substantially between domains and that different models trained on the same data are vulnerable to different attacks. 1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- Moral Stories: Situated Reasoning about Norms, Intents, Actions, and their ConsequencesDenis Emelin, Ronan Le Bras, Jena D. Hwang, Maxwell Forbes 等EMNLP 2021 · 被引用 83 次
- Nibbling at the Hard Core of Word Sense DisambiguationMarco Maru, Simone Conia, Michele Bevilacqua, Roberto NavigliACL 2022 · 被引用 22 次
- Contrastive Conditioning for Assessing Disambiguation in MT: A Case Study of Distilled BiasJannis Vamvas, Rico SennrichEMNLP 2021 · 被引用 12 次
- Reducing Sentiment Bias in Pre-trained Sentiment Classification via Adaptive Gumbel AttackJiachen Tian, Shizhan Chen, Xiaowang Zhang, Xin Wang 等AAAI 2023 · 被引用 5 次
- Revisiting Commonsense Reasoning in Machine Translation: Training, Evaluation and ChallengeXuebo Liu, Yutong Wang, Derek F. Wong, Runzhe Zhan 等ACL 2023 · 被引用 3 次
它引用的顶会 Paper1
相关 Paper
- Back Deduction Based Testing for Word Sense Disambiguation Ability of Machine Translation SystemsJun Wang, Yanhui Li, Xiang Huang, Lin Chen 等ISSTA 2023 · 被引用 4 次
- Do Context-Aware Translation Models Pay the Right Attention?Kayo Yin, Patrick Fernandes, Danish Pruthi, Aditi Chaudhary 等ACL 2021
- Imitation Attacks and Defenses for Black-box Machine Translation SystemsEric Wallace, Mitchell Stern, Dawn SongEMNLP 2020 · 被引用 63 次
- Crafting Adversarial Examples for Neural Machine TranslationXinze Zhang, Junzhe Zhang, Zhenhua Chen, Kun HeACL 2021
- DMDTEval: An Evaluation and Analysis of LLMs on Disambiguation in Multi-domain TranslationZhibo Man, Yuanmeng Chen, Yujie Zhang, Jinan XuEMNLP 2025
