DECOR: Improving Coherence in L2 English Writing with a Novel Benchmark for Incoherence Detection, Reasoning, and Rewriting
Xuanming Zhang, Anthony Diaz, Zixun Chen, Qingyang Wu, Kun Qian, Erik Voss, Zhou Yu
Abstract
Coherence in writing, an aspect that secondlanguage (L2) English learners often struggle with, is crucial in assessing L2 English writing. Existing automated writing evaluation systems primarily use basic surface linguistic features to detect coherence in writing. However, little effort has been made to correct the detected incoherence, which could significantly benefit L2 language learners seeking to improve their writing. To bridge this gap, we introduce DECOR, a novel benchmark that includes expert annotations for detecting incoherence in L2 English writing, identifying the underlying reasons, and rewriting the incoherent sentences. To our knowledge, DECOR is the first coherence assessment dataset specifically designed for improving L2 English writing, featuring pairs of original incoherent sentences alongside their expert-rewritten counterparts. Additionally, we fine-tuned models to automatically detect and rewrite incoherence in student essays. We find that incorporating specific reasons for incoherence during fine-tuning consistently improves the quality of the rewrites, achieving a result that is favored in both automatic and human evaluations. 1
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3baa35be-7f5d-4bcd-b31f-054ef991eef2Cited by top-tier papers1
Ask how each one uses itBuilds on5
- LIMA: Less Is More for AlignmentChunting Zhou, Pengfei Liu, Puxin Xu, Srinivasan Iyer et al.NeurIPS 2023 · 1,486 citations
- AlpacaFarm: A Simulation Framework for Methods that Learn from Human FeedbackYann Dubois, Chen Xuechen Li, Rohan Taori, Tianyi Zhang et al.NeurIPS 2023 · 948 citations
- LM-Critic: Language Models for Unsupervised Grammatical Error CorrectionMichihiro Yasunaga, Jure Leskovec, Percy LiangEMNLP 2021 · 29 citations
- Ensembling and Knowledge Distilling of Large Sequence Taggers for Grammatical Error CorrectionMaksym Tarnavskyi, Artem N. Chernodub, Kostiantyn OmelianchukACL 2022 · 28 citations
- COHESENTIA: A Novel Benchmark of Incremental versus Holistic Assessment of Coherence in Generated TextsAviya Maimon, Reut TsarfatyEMNLP 2023 · 2 citations
Related papers
- Centering-based Neural Coherence Modeling with Hierarchical Discourse SegmentsSungho Jeon, Michael StrubeEMNLP 2020 · 10 citations
- Towards Quantifiable Dialogue Coherence EvaluationZheng Ye, Liucun Lu, Lishan Huang, Liang Lin et al.ACL 2021
- DREsS: Dataset for Rubric-based Essay Scoring on EFL WritingHaneul Yoo, Jieun Han, So-Yeon Ahn, Alice OhACL 2025 · 13 citations
- A Multi-Task Dataset for Assessing Discourse Coherence in Chinese Essays: Structure, Theme, and Logic AnalysisHongyi Wu, Xinshu Shen, Man Lan, Shaoguang Mao et al.EMNLP 2023 · 4 citations
- Joint Modeling of Entities and Discourse Relations for Coherence AssessmentWei Liu, Michael StrubeEMNLP 2025
