DirectQE: Direct Pretraining for Machine Translation Quality Estimation
Qu Cui, Shujian Huang, Jiahuan Li, Xiang Geng, Zaixiang Zheng, Guoping Huang, Jiajun Chen
摘要
Machine Translation Quality Estimation (QE) is a task of predicting the quality of machine translations without relying on any reference. Recently, the predictor-estimator framework trains the predictor as a feature extractor, which leverages the extra parallel corpora without QE labels, achieving promising QE performance. However, we argue that there are gaps between the predictor and the estimator in both data quality and training objectives, which preclude QE models from benefiting from a large number of parallel corpora more directly. We propose a novel framework called DirectQE that provides a direct pretraining for QE tasks. In DirectQE, a generator is trained to produce pseudo data that is closer to the real QE data, and a detector is pretrained on these data with novel objectives that are akin to the QE task. Experiments on widely used benchmarks show that DirectQE outperforms existing methods, without using any pretraining models such as BERT. We also give extensive analyses showing how fixing the two gaps contributes to our improvements.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Cross-Lingual Cross-Modal Retrieval with Noise-Robust LearningYabing Wang, Jianfeng Dong, Tianxiang Liang, Minsong Zhang 等ACM MM 2022 · 被引用 26 次
- Denoising Pre-training for Machine Translation Quality Estimation with Curriculum LearningXiang Geng, Yu Zhang, Jiahuan Li, Shujian Huang 等AAAI 2023 · 被引用 11 次
- Self-Supervised Quality Estimation for Machine TranslationYuanhang Zheng, Zhixing Tan, Meng Zhang, Mieradilijiang Maimaiti 等EMNLP 2021 · 被引用 5 次
- Improved Pseudo Data for Machine Translation Quality Estimation with Constrained Beam SearchXiang Geng, Yu Zhang, Zhejian Lai, Shuaijie She 等EMNLP 2023 · 被引用 2 次
- Alleviating Distribution Shift in Synthetic Data for Machine Translation Quality EstimationXiang Geng, Zhejian Lai, Jiajun Chen, Hao Yang 等ACL 2025
它引用的顶会 Paper1
相关 Paper
- Bias Mitigation in Machine Translation Quality EstimationHanna Behnke, Marina Fomicheva, Lucia SpeciaACL 2022
- COMET: A Neural Framework for MT EvaluationRicardo Rei, Craig Stewart, Ana C. Farinha, Alon LavieEMNLP 2020 · 被引用 6 次
- PreQuEL: Quality Estimation of Machine Translation Outputs in AdvanceShachar Don-Yehiya, Leshem Choshen, Omri AbendEMNLP 2022 · 被引用 4 次
- Verdi: Quality Estimation and Error Detection for Bilingual CorporaMingjun Zhao, Haijiang Wu, Di Niu, Zixuan Wang 等WWW 2021 · 被引用 3 次
- Classification-based Quality Estimation: Small and Efficient Models for Real-world ApplicationsShuo Sun, Ahmed El-Kishky, Vishrav Chaudhary, James Cross 等EMNLP 2021 · 被引用 1 次
