Can Transformer be Too Compositional? Analysing Idiom Processing in Neural Machine Translation
Verna Dankers, Christopher G. Lucas, Ivan Titov
Abstract
Unlike literal expressions, idioms' meanings do not directly follow from their parts, posing a challenge for neural machine translation (NMT). NMT models are often unable to translate idioms accurately and over-generate compositional, literal translations. In this work, we investigate whether the non-compositionality of idioms is reflected in the mechanics of the dominant NMT model, Transformer, by analysing the hidden states and attention patterns for models with English as source language and one of seven European languages as target language. When Transformer emits a non-literal translation -i.e. identifies the expression as idiomatic -the encoder processes idioms more strongly as single lexical units compared to literal expressions. This manifests in idioms' parts being grouped through attention and in reduced interaction between idioms and their context. In the decoder's cross-attention, figurative inputs result in reduced attention on source-side tokens. These results suggest that Transformer's tendency to process idioms as compositional expressions contributes to literal translations of idioms.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 5924380a-3282-46b1-9b09-445d9246db2dCited by top-tier papers15
- LEACE: Perfect linear concept erasure in closed formNora Belrose, David Schneider-Joseph, Shauli Ravfogel, Ryan Cotterell et al.NeurIPS 2023 · 305 citations
- Translate Meanings, Not Just Words: IdiomKB's Role in Optimizing Idiomatic Translation with Language ModelsShuang Li, Jiangjie Chen, Siyu Yuan, Xinyi Wu et al.AAAI 2024 · 44 citations
- Divergences between Language Models and Human BrainsYuchen Zhou, Emmy Liu, Graham Neubig, Michael J. Tarr et al.NeurIPS 2024 · 8 citations
- Are representations built from the ground up? An empirical examination of local composition in language modelsEmmy Liu, Graham NeubigEMNLP 2022 · 5 citations
- Better Hit the Nail on the Head than Beat around the Bush: Removing Protected Attributes with a Single ProjectionPantea Haghighatkhah, Antske Fokkens, Pia Sommerauer, Bettina Speckmann et al.EMNLP 2022 · 2 citations
Builds on3
- Null It Out: Guarding Protected Attributes by Iterative Nullspace ProjectionShauli Ravfogel, Yanai Elazar, Hila Gonen, Michael Twiton et al.ACL 2020 · 25 citations
- The Paradox of the Compositionality of Natural Language: A Neural Machine Translation Case StudyVerna Dankers, Elia Bruni, Dieuwke HupkesACL 2022
- On Compositional Generalization of Neural Machine TranslationYafu Li, Yongjing Yin, Yulong Chen, Yue ZhangACL 2021
Related papers
- It's Not a Walk in the Park! Challenges of Idiom Translation in Speech-to-text SystemsIuliia Zaitova, Badr M. Abdullah, Wei Xue, Dietrich Klakow et al.ACL 2025 · 1 citation
- Crossing the Threshold: Idiomatic Machine Translation through Retrieval Augmentation and Loss WeightingEmmy Liu, Aditi Chaudhary, Graham NeubigEMNLP 2023 · 2 citations
- Characterizing Idioms: Conventionality and ContingencyMichaela Socolof, Jackie Chi Kit Cheung, Michael Wagner, Timothy J. O'DonnellACL 2022 · 9 citations
- Rethinking the Idiomaticity Decomposability Hypothesis: Evidence from Distributional LearningMaggie Mi, Golzar Atefi, Atsuki Yamaguchi, Felix A. Gers et al.ACL 2026
- Memorization or Reasoning? Exploring the Idiom Understanding of LLMsJisu Kim, Youngwoo Shin, Uiji Hwang, Jihun Choi et al.EMNLP 2025
