System Combination via Quality Estimation for Grammatical Error Correction
Muhammad Reza Qorib, Hwee Tou Ng
Abstract
Quality estimation models have been developed to assess the corrections made by grammatical error correction (GEC) models when the reference or gold-standard corrections are not available. An ideal quality estimator can be utilized to combine the outputs of multiple GEC systems by choosing the best subset of edits from the union of all edits proposed by the GEC base systems. However, we found that existing GEC quality estimation models are not good enough in differentiating good corrections from bad ones, resulting in a low F 0.5 score when used for system combination. In this paper, we propose GRECO 1 , a new state-of-the-art quality estimation model that gives a better estimate of the quality of a corrected sentence, as indicated by having a higher correlation to the F 0.5 score of a corrected sentence. It results in a combined GEC system with a higher F 0.5 score. We also propose three methods for utilizing GEC quality estimation models for system combination with varying generality: modelagnostic, model-agnostic with voting bias, and model-dependent method. The combined GEC system outperforms the state of the art on the CoNLL-2014 test set and the BEA-2019 test set, achieving the highest F 0.5 scores published to date.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ae5e91fa-5de4-4cf8-9005-c0d2086f307bCited by top-tier papers3
- Leveraging What's Overfixed: Post-Correction via LLM Grammatical Error OvercorrectionTaehee Park, Heejin Do, Gary LeeEMNLP 2025 · 1 citation
- Enhancing Text Editing for Grammatical Error Correction: Arabic as a Case StudyBashar Alhafni, Nizar HabashACL 2025
- JELV: A Judge of Edit-Level Validity for Evaluation and Automated Reference Expansion in Grammatical Error CorrectionYuhao Zhan, Yuqing Zhang, Jing Yuan, Qixiang Ma et al.AAAI 2026
Builds on3
- DeBERTaV3: Improving DeBERTa using ELECTRA-Style Pre-Training with Gradient-Disentangled Embedding SharingPengcheng He, Jianfeng Gao, Weizhu ChenICLR 2023 · 394 citations
- Ensembling and Knowledge Distilling of Large Sequence Taggers for Grammatical Error CorrectionMaksym Tarnavskyi, Artem N. Chernodub, Kostiantyn OmelianchukACL 2022 · 28 citations
- Improved grammatical error correction by ranking elementary editsAlexey SorokinEMNLP 2022 · 10 citations
Related papers
- CLEME: Debiasing Multi-reference Evaluation for Grammatical Error CorrectionJingheng Ye, Yinghui Li, Qingyu Zhou, Yangning Li et al.EMNLP 2023 · 5 citations
- Multi-Class Grammatical Error Detection for Correction: A Tale of Two SystemsZheng Yuan, Shiva Taslimipoor, Christopher Davis, Christopher BryantEMNLP 2021 · 28 citations
- Revisiting Grammatical Error Correction Evaluation and BeyondPeiyuan Gong, Xuebo Liu, Heyan Huang, Min ZhangEMNLP 2022 · 11 citations
- Grammatical Error Correction in Low Error Density Domains: A New Benchmark and AnalysesSimon Flachs, Ophélie Lacroix, Helen Yannakoudakis, Marek Rei et al.EMNLP 2020
- RobustGEC: Robust Grammatical Error Correction Against Subtle Context PerturbationYue Zhang, Leyang Cui, Enbo Zhao, Wei Bi et al.EMNLP 2023 · 1 citation
