Mask-Align: Self-Supervised Neural Word Alignment
Chi Chen, Maosong Sun, Yang Liu
Abstract
Word alignment, which aims to align translationally equivalent words between source and target sentences, plays an important role in many natural language processing tasks. Current unsupervised neural alignment methods focus on inducing alignments from neural machine translation models, which does not leverage the full context in the target sequence. In this paper, we propose MASK-ALIGN, a selfsupervised word alignment model that takes advantage of the full context on the target side. Our model parallelly masks out each target token and predicts it conditioned on both source and the remaining target tokens. This two-step process is based on the assumption that the source token contributing most to recovering the masked target token should be aligned. We also introduce an attention variant called leaky attention, which alleviates the problem of high cross-attention weights on specific tokens such as periods. Experiments on four language pairs show that our model outperforms previous unsupervised neural aligners and obtains new state-of-the-art results. 1 * Corresponding author 1 Code can be found at https://github.com/THUNLP-MT/ Mask-Align .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 08bbf16b-ac88-48ec-a0a2-0960cc7dead3Cited by top-tier papers6
- What learning algorithm is in-context learning? Investigations with linear modelsEkin Akyürek, Dale Schuurmans, Jacob Andreas, Tengyu Ma et al.ICLR 2023 · 85 citations
- Prompting Neural Machine Translation with Translation MemoriesAbudurexiti Reheman, Tao Zhou, Yingfeng Luo, Di Yang et al.AAAI 2023 · 11 citations
- Cross-Align: Modeling Deep Cross-lingual Interactions for Word AlignmentSiyu Lai, Zhen Yang, Fandong Meng, Yufeng Chen et al.EMNLP 2022 · 6 citations
- Self-Supervised Quality Estimation for Machine TranslationYuanhang Zheng, Zhixing Tan, Meng Zhang, Mieradilijiang Maimaiti et al.EMNLP 2021 · 5 citations
- BinaryAlign: Word Alignment as Binary Sequence LabelingGaetan Latouche, Marc-André Carbonneau, Benjamin SwansonACL 2024 · 2 citations
Builds on3
- Non-autoregressive Machine Translation with Disentangled Context TransformerJungo Kasai, James Cross, Marjan Ghazvininejad, Jiatao GuICML 2020 · 113 citations
- Accurate Word Alignment Induction from Neural Machine TranslationYun Chen, Yang Liu, Guanhua Chen, Xin Jiang et al.EMNLP 2020 · 56 citations
- End-to-End Neural Word Alignment Outperforms GIZA++Thomas Zenkel, Joern Wuebker, John DeNeroACL 2020 · 2 citations
Related papers
- A Bidirectional Transformer Based Alignment Model for Unsupervised Word AlignmentJingyi Zhang, Josef van GenabithACL 2021
- Self-supervised Bilingual Syntactic Alignment for Neural Machine TranslationTianfu Zhang, Heyan Huang, Chong Feng, Longbing CaoAAAI 2021 · 7 citations
- Improving Pretrained Cross-Lingual Language Models via Self-Labeled Word AlignmentZewen Chi, Li Dong, Bo Zheng, Shaohan Huang et al.ACL 2021
- A Supervised Word Alignment Method based on Cross-Language Span Prediction using Multilingual BERTMasaaki Nagata, Katsuki Chousa, Masaaki NishinoEMNLP 2020 · 38 citations
- SenseBERT: Driving Some Sense into BERTYoav Levine, Barak Lenz, Or Dagan, Ori Ram et al.ACL 2020 · 27 citations
