GuoFeng: A Benchmark for Zero Pronoun Recovery and Translation
Mingzhou Xu, Longyue Wang, Derek F. Wong, Hongye Liu, Linfeng Song, Lidia S. Chao, Shuming Shi, Zhaopeng Tu
Abstract
The phenomenon of zero pronoun (ZP) has attracted increasing interest in the machine translation (MT) community due to its importance and difficulty. However, previous studies generally evaluate the quality of translating ZPs with BLEU scores on MT testsets, which are not expressive or sensitive enough for accurate assessment. To bridge the data and evaluation gaps, we propose a benchmark testset for target evaluation on Chinese-English ZP translation. The human-annotated testset covers five challenging genres, which reveal different characteristics of ZPs for comprehensive evaluation. We systematically revisit eight advanced models on ZP translation and identify current challenges for future exploration. We release data, code, models and annotation guidelines, which we hope can significantly promote research in this field. 1 * Mingzhou Xu and Longyue Wang contributed equally to this work. Work was done when Mingzhou Xu and Hongye Liu were interning at Tencent AI Lab. 1 https://github.com/longyuewangdcu/ mZPRT . Inp. 黄娟 Huang Juan, female, associate professor. Mainly teach the course Business English. Inp. A: 菲比 很 想 买 台 电视 。 B: 乔伊 不 让 (她) 买 (它) ?
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 02510546-a4bc-4fd9-8a4b-a705be7b285fCited by top-tier papers5
- Document-Level Machine Translation with Large Language ModelsLongyue Wang, Chenyang Lyu, Tianbo Ji, Zhirui Zhang et al.EMNLP 2023 · 129 citations
- A Survey on Zero Pronoun TranslationLongyue Wang, Siyou Liu, Mingzhou Xu, Linfeng Song et al.ACL 2023 · 5 citations
- LiTransProQA: An LLM-based Literary Translation Evaluation Metric with Professional Question AnsweringRan Zhang, Wei Zhao, Lieve Macken, Steffen EgerEMNLP 2025 · 2 citations
- PROSE: A Pronoun Omission Solution for Chinese-English Spoken Language TranslationKe Wang, Xiutian Zhao, Yanghui Li, Wei PengEMNLP 2023 · 2 citations
- Who Wrote This Line? Evaluating the Detection of LLM-Generated Classical Chinese PoetryJiang Li, Tian Lan, Shanshan Wang, Zdongxing et al.ACL 2026
Builds on4
- Coupling Context Modeling with Zero Pronoun Recovering for Document-Level Natural Language GenerationXin Tan, Longyin Zhang, Guodong ZhouEMNLP 2021 · 6 citations
- COMET: A Neural Framework for MT EvaluationRicardo Rei, Craig Stewart, Ana C. Farinha, Alon LavieEMNLP 2020 · 6 citations
- UniTE: Unified Translation EvaluationYu Wan, Dayiheng Liu, Baosong Yang, Haibo Zhang et al.ACL 2022
- Document Graph for Neural Machine TranslationMingzhou Xu, Liangyou Li, Derek F. Wong, Qun Liu et al.EMNLP 2021
Related papers
- MT-GenEval: A Counterfactual and Contextual Dataset for Evaluating Gender Accuracy in Machine TranslationAnna Currey, Maria Nadejde, Raghavendra Reddy Pappagari, Mia Mayer et al.EMNLP 2022 · 22 citations
- DiscoX: Benchmarking Discourse-Level Translation in Expert DomainsXiying ZHAO, Zhoufutu Wen, Zhixuan Chen, Jingzhe Ding et al.ICLR 2026 · 2 citations
- Languages Still Left Behind: Toward a Better Multilingual Machine Translation BenchmarkChihiro Taguchi, Seng Mai, Keita Kurabe, Yusuke Sakai et al.EMNLP 2025
- Automatic Machine Translation Evaluation in Many Languages via Zero-Shot ParaphrasingBrian Thompson, Matt PostEMNLP 2020 · 7 citations
- Bilingual Zero-Shot Stance DetectionChenye Zhao, Cornelia CarageaACL 2025 · 1 citation
