On the Importance of Word Order Information in Cross-lingual Sequence Labeling
Zihan Liu, Genta Indra Winata, Samuel Cahyawijaya, Andrea Madotto, Zhaojiang Lin, Pascale Fung
摘要
Cross-lingual models trained on source language tasks possess the capability to directly transfer to target languages. However, since word order variances generally exist in different languages, cross-lingual models that overfit into the word order of the source language could have sub-optimal performance in target languages. In this paper, we hypothesize that reducing the word order information fitted into the models can improve the adaptation performance in target languages. To verify this hypothesis, we introduce several methods to make models encode less word order information of the source language and test them based on cross-lingual word embeddings and the pre-trained multilingual model. Experimental results on three sequence labeling tasks (i.e., part-of-speech tagging, named entity recognition and slot filling tasks) show that reducing word order information injected into the model can achieve better zero-shot cross-lingual performance. Further analysis illustrates that fitting excessive or insufficient word order information into the model results in inferior cross-lingual performance. Moreover, our proposed methods can also be applied to strong cross-lingual models and further improve their performance.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- The Impact of Positional Encodings on Multilingual CompressionVinit Ravishankar, Anders SøgaardEMNLP 2021 · 被引用 7 次
- Implicit Word Reordering with Knowledge Distillation for Cross-Lingual Dependency ParsingZhuoran Li, Chunming Hu, Junfan Chen, Zhijun Chen 等AAAI 2025 · 被引用 1 次
- Argument Mining in Data Scarce Settings: Cross-lingual Transfer and Few-shot TechniquesAnar Yeginbergen, Maite Oronoz, Rodrigo AgerriACL 2024
它引用的顶会 Paper4
- Unsupervised Cross-lingual Representation Learning at ScaleAlexis Conneau, Kartikay Khandelwal, Naman Goyal, Vishrav Chaudhary 等ACL 2020 · 被引用 539 次
- XGLUE: A New Benchmark Datasetfor Cross-lingual Pre-training, Understanding and GenerationYaobo Liang, Nan Duan, Yeyun Gong, Ning Wu 等EMNLP 2020 · 被引用 232 次
- Attention-Informed Mixed-Language Training for Zero-Shot Cross-Lingual Task-Oriented Dialogue SystemsZihan Liu, Genta Indra Winata, Zhaojiang Lin, Peng Xu 等AAAI 2020 · 被引用 105 次
- Cross-lingual Spoken Language Understanding with Regularized Representation AlignmentZihan Liu, Genta Indra Winata, Peng Xu, Zhaojiang Lin 等EMNLP 2020 · 被引用 16 次
相关 Paper
- Word Reordering for Zero-shot Cross-lingual Structured PredictionTao Ji, Yong Jiang, Tao Wang, Zhongqiang Huang 等EMNLP 2021 · 被引用 3 次
- Make the Best of Cross-lingual Transfer: Evidence from POS Tagging with over 100 LanguagesWietse de Vries, Martijn Wieling, Malvina NissimACL 2022 · 被引用 63 次
- A Comparison of Architectures and Pretraining Methods for Contextualized Multilingual Word EmbeddingsNiels van der Heijden, Samira Abnar, Ekaterina ShutovaAAAI 2020 · 被引用 16 次
- Evaluating morphological typology in zero-shot cross-lingual transferAntonio Martínez-García, Toni Badia, Jeremy BarnesACL 2021
- Multilingual Alignment of Contextual Word RepresentationsSteven Cao, Nikita Kitaev, Dan KleinICLR 2020 · 被引用 211 次
