GuoFeng: A Benchmark for Zero Pronoun Recovery and Translation
Mingzhou Xu, Longyue Wang, Derek F. Wong, Hongye Liu, Linfeng Song, Lidia S. Chao, Shuming Shi, Zhaopeng Tu
摘要
The phenomenon of zero pronoun (ZP) has attracted increasing interest in the machine translation (MT) community due to its importance and difficulty. However, previous studies generally evaluate the quality of translating ZPs with BLEU scores on MT testsets, which are not expressive or sensitive enough for accurate assessment. To bridge the data and evaluation gaps, we propose a benchmark testset for target evaluation on Chinese-English ZP translation. The human-annotated testset covers five challenging genres, which reveal different characteristics of ZPs for comprehensive evaluation. We systematically revisit eight advanced models on ZP translation and identify current challenges for future exploration. We release data, code, models and annotation guidelines, which we hope can significantly promote research in this field. 1 * Mingzhou Xu and Longyue Wang contributed equally to this work. Work was done when Mingzhou Xu and Hongye Liu were interning at Tencent AI Lab. 1 https://github.com/longyuewangdcu/ mZPRT . Inp. 黄娟 Huang Juan, female, associate professor. Mainly teach the course Business English. Inp. A: 菲比 很 想 买 台 电视 。 B: 乔伊 不 让 (她) 买 (它) ?
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Document-Level Machine Translation with Large Language ModelsLongyue Wang, Chenyang Lyu, Tianbo Ji, Zhirui Zhang 等EMNLP 2023 · 被引用 129 次
- A Survey on Zero Pronoun TranslationLongyue Wang, Siyou Liu, Mingzhou Xu, Linfeng Song 等ACL 2023 · 被引用 5 次
- LiTransProQA: An LLM-based Literary Translation Evaluation Metric with Professional Question AnsweringRan Zhang, Wei Zhao, Lieve Macken, Steffen EgerEMNLP 2025 · 被引用 2 次
- PROSE: A Pronoun Omission Solution for Chinese-English Spoken Language TranslationKe Wang, Xiutian Zhao, Yanghui Li, Wei PengEMNLP 2023 · 被引用 2 次
- Who Wrote This Line? Evaluating the Detection of LLM-Generated Classical Chinese PoetryJiang Li, Tian Lan, Shanshan Wang, Zdongxing 等ACL 2026
它引用的顶会 Paper4
- Coupling Context Modeling with Zero Pronoun Recovering for Document-Level Natural Language GenerationXin Tan, Longyin Zhang, Guodong ZhouEMNLP 2021 · 被引用 6 次
- COMET: A Neural Framework for MT EvaluationRicardo Rei, Craig Stewart, Ana C. Farinha, Alon LavieEMNLP 2020 · 被引用 6 次
- UniTE: Unified Translation EvaluationYu Wan, Dayiheng Liu, Baosong Yang, Haibo Zhang 等ACL 2022
- Document Graph for Neural Machine TranslationMingzhou Xu, Liangyou Li, Derek F. Wong, Qun Liu 等EMNLP 2021
相关 Paper
- MT-GenEval: A Counterfactual and Contextual Dataset for Evaluating Gender Accuracy in Machine TranslationAnna Currey, Maria Nadejde, Raghavendra Reddy Pappagari, Mia Mayer 等EMNLP 2022 · 被引用 22 次
- DiscoX: Benchmarking Discourse-Level Translation in Expert DomainsXiying ZHAO, Zhoufutu Wen, Zhixuan Chen, Jingzhe Ding 等ICLR 2026 · 被引用 2 次
- Languages Still Left Behind: Toward a Better Multilingual Machine Translation BenchmarkChihiro Taguchi, Seng Mai, Keita Kurabe, Yusuke Sakai 等EMNLP 2025
- Automatic Machine Translation Evaluation in Many Languages via Zero-Shot ParaphrasingBrian Thompson, Matt PostEMNLP 2020 · 被引用 7 次
- Bilingual Zero-Shot Stance DetectionChenye Zhao, Cornelia CarageaACL 2025 · 被引用 1 次
