From Good to Best: Two-Stage Training for Cross-Lingual Machine Reading Comprehension
Nuo Chen, Linjun Shou, Ming Gong, Jian Pei
摘要
Cross-lingual Machine Reading Comprehension (xMRC) is a challenging task due to the lack of training data in low-resource languages. Recent approaches use training data only in a resource-rich language (such as English) to fine-tune large-scale cross-lingual pre-trained language models, which transfer knowledge from resource-rich languages (source) to low-resource languages (target). Due to the big difference between languages, the model fine-tuned only by the source language may not perform well for target languages. In our study, we make an interesting observation that while the top 1 result predicted by the previous approaches may often fail to hit the ground-truth answer, there are still good chances for the correct answer to be contained in the set of top k predicted results. Intuitively, the previous approaches have empowered the model certain level of capability to roughly distinguish good answers from bad ones. However, without sufficient training data, it is not powerful enough to capture the nuances between the accurate answer and those approximate ones. Based on this observation, we develop a two-stage approach to enhance the model performance. The first stage targets at recall; we design a hard-learning (HL) algorithm to maximize the likelihood that the top k predictions contain the accurate answer. The second stage focuses on precision, where an answer-aware contrastive learning (AA-CL) mechanism is developed to learn the minute difference between the accurate answer and other candidates. Extensive experiments show that our model significantly outperforms strong baselines on two cross-lingual MRC benchmark datasets.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Alleviating Over-smoothing for Unsupervised Sentence RepresentationNuo Chen, Linjun Shou, Jian Pei, Ming Gong 等ACL 2023 · 被引用 10 次
- Enhancing Pre-Trained Generative Language Models with Question Attended Span Extraction on Machine Reading ComprehensionLin Ai, Zheng Hui, Zizhou Liu, Julia HirschbergEMNLP 2024 · 被引用 4 次
它引用的顶会 Paper6
- SimCSE: Simple Contrastive Learning of Sentence EmbeddingsTianyu Gao, Xingcheng Yao, Danqi ChenEMNLP 2021 · 被引用 2,496 次
- Unsupervised Cross-lingual Representation Learning at ScaleAlexis Conneau, Kartikay Khandelwal, Naman Goyal, Vishrav Chaudhary 等ACL 2020 · 被引用 539 次
- XGLUE: A New Benchmark Datasetfor Cross-lingual Pre-training, Understanding and GenerationYaobo Liang, Nan Duan, Yeyun Gong, Ning Wu 等EMNLP 2020 · 被引用 232 次
- MLQA: Evaluating Cross-lingual Extractive Question AnsweringPatrick Lewis, Barlas Oguz, Ruty Rinott, Sebastian Riedel 等ACL 2020 · 被引用 52 次
- Enhancing Answer Boundary Detection for Multilingual Machine Reading ComprehensionFei Yuan, Linjun Shou, Xuanyu Bai, Ming Gong 等ACL 2020 · 被引用 21 次
相关 Paper
- Learning Disentangled Semantic Representations for Zero-Shot Cross-Lingual Transfer in Multilingual Machine Reading ComprehensionLinjuan Wu, Shaojuan Wu, Xiaowang Zhang, Deyi Xiong 等ACL 2022 · 被引用 18 次
- KECP: Knowledge Enhanced Contrastive Prompting for Few-shot Extractive Question AnsweringJianing Wang, Chengyu Wang, Minghui Qiu, Qiuhui Shi 等EMNLP 2022 · 被引用 16 次
- MTMS: Multi-teacher Multi-stage Knowledge Distillation for Reasoning-Based Machine Reading ComprehensionZhuo Zhao, Zhiwen Xie, Guangyou Zhou, Jimmy Xiangji HuangSIGIR 2024 · 被引用 10 次
- Representation and Labeling Gap Bridging for Cross-lingual Named Entity RecognitionXinghua Zhang, Bowen Yu, Jiangxia Cao, Quangang Li 等SIGIR 2023 · 被引用 5 次
- Improving Word Translation via Two-Stage Contrastive LearningYaoyiran Li, Fangyu Liu, Nigel Collier, Anna Korhonen 等ACL 2022 · 被引用 32 次
