Syntax-augmented Multilingual BERT for Cross-lingual Transfer
Wasi Uddin Ahmad, Haoran Li, Kai-Wei Chang, Yashar Mehdad
Abstract
In recent years, we have seen a colossal effort in pre-training multilingual text encoders using large-scale corpora in many languages to facilitate cross-lingual transfer learning. However, due to typological differences across languages, the cross-lingual transfer is challenging. Nevertheless, language syntax, e.g., syntactic dependencies, can bridge the typological gap. Previous works have shown that pretrained multilingual encoders, such as mBERT (Devlin et al., 2019) , capture language syntax, helping cross-lingual transfer. This work shows that explicitly providing language syntax and training mBERT using an auxiliary objective to encode the universal dependency tree structure helps cross-lingual transfer. We perform rigorous experiments on four NLP tasks, including text classification, question answering, named entity recognition, and taskoriented semantic parsing. The experiment results show that syntax-augmented mBERT improves cross-lingual transfer on popular benchmarks, such as PAWS-X and MLQA, by 1.4 and 1.6 points on average across all languages. In the generalized transfer setting, the performance boosted significantly, with 3.9 and 3.1 points on average in PAWS-X and MLQA.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 211e7db3-d135-4230-9719-3afbb6f52db1Cited by top-tier papers5
- Syntax and Domain Aware Model for Unsupervised Program TranslationFang Liu, Jia Li, Li ZhangICSE 2023 · 25 citations
- Multilingual Pre-training with Universal Dependency LearningKailai Sun, Zuchao Li, Hai ZhaoNeurIPS 2021 · 11 citations
- Struct-XLM: A Structure Discovery Multilingual Language Model for Enhancing Cross-lingual Transfer through Reinforcement LearningLinjuan Wu, Weiming LuEMNLP 2023
- On the Continued Value of Universal Dependencies in the Era of Large Language ModelsWenxi Li, Jingyu PengACL 2026
- An Unsupervised, Geometric and Syntax-aware Quantification of PolysemyAnmol Goel, Charu Sharma, Ponnurangam KumaraguruEMNLP 2022
Builds on9
- Unsupervised Cross-lingual Representation Learning at ScaleAlexis Conneau, Kartikay Khandelwal, Naman Goyal, Vishrav Chaudhary et al.ACL 2020 · 539 citations
- Cross-Lingual Ability of Multilingual BERT: An Empirical StudyKarthikeyan K, Zihan Wang, Stephen Mayhew, Dan RothICLR 2020 · 378 citations
- XGLUE: A New Benchmark Datasetfor Cross-lingual Pre-training, Understanding and GenerationYaobo Liang, Nan Duan, Yeyun Gong, Ning Wu et al.EMNLP 2020 · 232 citations
- SG-Net: Syntax-Guided Machine Reading ComprehensionZhuosheng Zhang, Yuwei Wu, Junru Zhou, Sufeng Duan et al.AAAI 2020 · 192 citations
- GATE: Graph Attention Transformer Encoder for Cross-lingual Relation and Event ExtractionWasi Uddin Ahmad, Nanyun Peng, Kai-Wei ChangAAAI 2021 · 113 citations
Related papers
- COSY: COunterfactual SYntax for Cross-Lingual UnderstandingSicheng Yu, Hao Zhang, Yulei Niu, Qianru Sun et al.ACL 2021
- XLM-K: Improving Cross-Lingual Language Model Pre-training with Multilingual KnowledgeXiaoze Jiang, Yaobo Liang, Weizhu Chen, Nan DuanAAAI 2022 · 31 citations
- Syntax-Enhanced Pre-trained ModelZenan Xu, Daya Guo, Duyu Tang, Qinliang Su et al.ACL 2021
- Finding Universal Grammatical Relations in Multilingual BERTEthan A. Chi, John Hewitt, Christopher D. ManningACL 2020 · 7 citations
- Cross-Linguistic Syntactic Difference in Multilingual BERT: How Good is It and How Does It Affect Transfer?Ningyu Xu, Tao Gui, Ruotian Ma, Qi Zhang et al.EMNLP 2022 · 4 citations
