GWLAN: General Word-Level AutocompletioN for Computer-Aided Translation
Huayang Li, Lemao Liu, Guoping Huang, Shuming Shi
摘要
Computer-aided translation (CAT), the use of software to assist a human translator in the translation process, has been proven to be useful in enhancing the productivity of human translators. Autocompletion, which suggests translation results according to the text pieces provided by human translators, is a core function of CAT. There are two limitations in previous research in this line. First, most research works on this topic focus on sentence-level autocompletion (i.e., generating the whole translation as a sentence based on human input), but word-level autocompletion is under-explored so far. Second, almost no public benchmarks are available for the autocompletion task of CAT. This might be among the reasons why research progress in CAT is much slower compared to automatic MT. In this paper, we propose the task of general word-level autocompletion (GWLAN) from a real-world CAT scenario, and construct the first public benchmark 1 to facilitate research in this topic. In addition, we propose an effective method for GWLAN and compare it with several strong baselines. Experiments demonstrate that our proposed method can give significantly more accurate predictions than the baseline methods on our benchmark datasets.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Skimming, Locating, then Perusing: A Human-Like Framework for Natural Language Video LocalizationDaizong Liu, Wei HuACM MM 2022 · 被引用 33 次
- BiTIIMT: A Bilingual Text-infilling Method for Interactive Machine TranslationYanling Xiao, Lemao Liu, Guoping Huang, Qu Cui 等ACL 2022 · 被引用 21 次
- WeTS: A Benchmark for Translation SuggestionZhen Yang, Fandong Meng, Yingxue Zhang, Ernan Li 等EMNLP 2022 · 被引用 5 次
- Rethinking Word-Level Auto-Completion in Computer-Aided TranslationXingyu Chen, Lemao Liu, Guoping Huang, Zhirui Zhang 等EMNLP 2023 · 被引用 3 次
它引用的顶会 Paper1
相关 Paper
- Cross-lingual neural fuzzy matching for exploiting target-language monolingual corpora in computer-aided translationMiquel Esplà-Gomis, Víctor M. Sánchez-Cartagena, Juan Antonio Pérez-Ortiz, Felipe Sánchez-MartínezEMNLP 2022 · 被引用 2 次
- Investigating the Helpfulness of Word-Level Quality Estimation for Post-Editing Machine Translation OutputRaksha Shenoy, Nico Herbig, Antonio Krüger, Josef van GenabithEMNLP 2021 · 被引用 3 次
- Improving Machine Translation Systems via Isotopic ReplacementZeyu Sun, Jie M. Zhang, Yingfei Xiong, Mark Harman 等ICSE 2022 · 被引用 42 次
- NAT4AT: Using Non-Autoregressive Translation Makes Autoregressive Translation Faster and BetterHuanran Zheng, Wei Zhu, Xiaoling WangWWW 2024 · 被引用 13 次
- Computer-Aided Tagging on Wikimedia Commons: Designing for Human–AI Collaboration in Open Knowledge WorkYihan Yu, David W. McDonaldCSCW 2026
