Do Context-Aware Translation Models Pay the Right Attention?
Kayo Yin, Patrick Fernandes, Danish Pruthi, Aditi Chaudhary, André F. T. Martins, Graham Neubig
摘要
Context-aware machine translation models are designed to leverage contextual information, but often fail to do so. As a result, they inaccurately disambiguate pronouns and polysemous words that require context for resolution. In this paper, we ask several questions: What contexts do human translators use to resolve ambiguous words? Are models paying large amounts of attention to the same context? What if we explicitly train them to do so? To answer these questions, we introduce SCAT (Supporting Context for Ambiguous Translations), a new English-French dataset comprising supporting context words for 14K translations that professional translators found useful for pronoun disambiguation. Using SCAT, we perform an in-depth analysis of the context used to disambiguate, examining positional and lexical characteristics of the supporting words. Furthermore, we measure the degree of alignment between the model's attention scores and the supporting context from SCAT, and apply a guided attention strategy to encourage agreement between the two. 1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Interpreting Language Models with Contrastive ExplanationsKayo Yin, Graham NeubigEMNLP 2022 · 被引用 32 次
- When Does Translation Require Context? A Data-driven, Multilingual ExplorationPatrick Fernandes, Kayo Yin, Emmy Liu, André F. T. Martins 等ACL 2023 · 被引用 10 次
- Quantifying the Plausibility of Context Reliance in Neural Machine TranslationGabriele Sarti, Grzegorz Chrupala, Malvina Nissim, Arianna BisazzaICLR 2024 · 被引用 8 次
- XplainLLM: A Knowledge-Augmented Dataset for Reliable Grounded Explanations in LLMsZichen Chen, Jianda Chen, Ambuj K. Singh, Misha SraEMNLP 2024 · 被引用 5 次
- A Tale of Pronouns: Interpretability Informs Gender Bias Mitigation for Fairer Instruction-Tuned Machine TranslationGiuseppe Attanasio, Flor Miriam Plaza del Arco, Debora Nozza, Anne LauscherEMNLP 2023 · 被引用 3 次
它引用的顶会 Paper3
- Less is More: Attention Supervision with Counterfactuals for Text ClassificationSeungtaek Choi, Haeju Park, Jinyoung Yeo, Seung-won HwangEMNLP 2020 · 被引用 16 次
- COMET: A Neural Framework for MT EvaluationRicardo Rei, Craig Stewart, Ana C. Farinha, Alon LavieEMNLP 2020 · 被引用 6 次
- Detecting Word Sense Disambiguation Biases in Machine Translation for Model-Agnostic Adversarial AttacksDenis Emelin, Ivan Titov, Rico SennrichEMNLP 2020 · 被引用 3 次
相关 Paper
- You Are What You Train: Effects of Data Composition on Training Context-aware Machine Translation ModelsPawel Maka, Yusuf Can Semerci, Jan Scholtes, Gerasimos SpanakisEMNLP 2025
- Seeing Through Ambiguity: Effective Video-guided Machine Translation via Chaotic Fusion and Causally Aligned Spatio-temporal AttentionJiawei Zheng, Feiyan Liu, Xiaoli WangACM MM 2025
- Video-Helpful Multimodal Machine TranslationYihang Li, Shuichiro Shimizu, Chenhui Chu, Sadao Kurohashi 等EMNLP 2023 · 被引用 1 次
- Visual Agreement Regularized Training for Multi-Modal Machine TranslationPengcheng Yang, Boxing Chen, Pei Zhang, Xu SunAAAI 2020 · 被引用 34 次
- DMDTEval: An Evaluation and Analysis of LLMs on Disambiguation in Multi-domain TranslationZhibo Man, Yuanmeng Chen, Yujie Zhang, Jinan XuEMNLP 2025
