EXCGEC: A Benchmark for Edit-Wise Explainable Chinese Grammatical Error Correction
Jingheng Ye, Shang Qin, Yinghui Li, Xuxin Cheng, Libo Qin, Hai-Tao Zheng, Ying Shen, Peng Xing, Zishan Xu, Guo Cheng, Wenhao Jiang
Abstract
Existing studies explore the explainability of Grammatical Error Correction (GEC) in a limited scenario, where they ignore the interaction between corrections and explanations and have not established a corresponding comprehensive benchmark. To bridge the gap, this paper first introduces the task of EXplainable GEC (EXGEC), which focuses on the integral role of correction and explanation tasks. To facilitate the task, we propose EXCGEC, a tailored benchmark for Chinese EXGEC consisting of 8,216 explanation-augmented samples featuring the design of hybrid edit-wise explanations. We then benchmark several series of LLMs in multi-task learning settings, including post-explaining and pre-explaining. To promote the development of the task, we also build a comprehensive evaluation suite by leveraging existing automatic metrics and conducting human evaluation experiments to demonstrate the human consistency of the automatic metrics for free-text explanations. Our experiments reveal the effectiveness of evaluating free-text explanations using traditional metrics like METEOR and ROUGE, and the inferior performance of multi-task models compared to the pipeline solution, indicating its challenges to establish positive effects in learning both tasks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b4766858-72f0-4b3d-b7b1-57935b1548ffCited by top-tier papers1
Ask how each one uses itBuilds on5
- LLM-powered Data Augmentation for Enhanced Cross-lingual PerformanceChenxi Whitehouse, Monojit Choudhury, Alham Fikri AjiEMNLP 2023 · 53 citations
- Unifying Model Explainability and Robustness for Joint Text Classification and Rationale ExtractionDongfang Li, Baotian Hu, Qingcai Chen, Tujie Xu et al.AAAI 2022 · 16 citations
- Enhancing Grammatical Error Correction Systems with ExplanationsYuejiao Fei, Leyang Cui, Sen Yang, Wai Lam et al.ACL 2023 · 13 citations
- CLEME: Debiasing Multi-reference Evaluation for Grammatical Error CorrectionJingheng Ye, Yinghui Li, Qingyu Zhou, Yangning Li et al.EMNLP 2023 · 5 citations
- Interpretability for Language Learners Using Example-Based Grammatical Error CorrectionMasahiro Kaneko, Sho Takase, Ayana Niwa, Naoaki OkazakiACL 2022
Related papers
- CL²GEC: A Multi-Discipline Benchmark for Continual Learning in Chinese Literature Grammatical Error CorrectionShang Qin, Jingheng Ye, Yinghui Li, Hai-Tao Zheng et al.ACL 2026
- Detection-Correction Structure via General Language Model for Grammatical Error CorrectionWei Li, Houfeng WangACL 2024 · 9 citations
- RobustGEC: Robust Grammatical Error Correction Against Subtle Context PerturbationYue Zhang, Leyang Cui, Enbo Zhao, Wei Bi et al.EMNLP 2023 · 1 citation
- Intuitive Thinking: Expanding Large Language Models' Thinking for Rapid Decision-Making on Candidate Corrections in Chinese Grammar Error CorrectionLintao Long, Ruizhang Huang, Ruina Bai, Yongbin Qin et al.AAAI 2026
- A Training-free LLM-based Approach to General Chinese Character Error CorrectionHouquan Zhou, Bo Zhang, Zhenghua Li, Ming Yan et al.ACL 2025
