DREsS: Dataset for Rubric-based Essay Scoring on EFL Writing
Haneul Yoo, Jieun Han, So-Yeon Ahn, Alice Oh
摘要
Automated essay scoring (AES) is a useful tool in English as a Foreign Language (EFL) writing education, offering real-time essay scores for students and instructors. However, previous AES models were trained on essays and scores irrelevant to the practical scenarios of EFL writing education and usually provided a single holistic score due to the lack of appropriate datasets. In this paper, we release DREsS, a large-scale, standard dataset for rubric-based automated essay scoring with 48.9K samples in total. DREsS comprises three sub-datasets: DREsS New , DREsS Std. , and DREsS CASE . We collect DREsS New , a real-classroom dataset with 2.3K essays authored by EFL undergraduate students and scored by English education experts. We also standardize existing rubricbased essay scoring datasets as DREsS Std. . We suggest CASE, a corruption-based augmentation strategy for essays, which generates 40.1K synthetic samples of DREsS CASE and improves the baseline results by 45.44%. DREsS will enable further research to provide a more accurate and practical AES system for EFL writing education. 1 1. DREsS_New (2,279 samples) EFL classroom data: 1) Student-written essays 2) Rubric-based scores assessed by instructors 2. DREsS_Std. (6,515 samples) Unified AES datasets with standardized rubrics under professional consultation Corruption 3. DREsS_CASE (40,185 samples) Synthetic essay samples generated by CASE, our proposed augmentation strategy
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Authorship Drift: How Self-Efficacy and Trust Evolve During LLM-Assisted WritingYeon Su Park, Nadia Azzahra Putri Arvi, Seoyoung Kim, Juho KimCHI 2026 · 被引用 3 次
- STRICTA: Structured Reasoning in Critical Text Assessment for Peer Review and BeyondNils Dycke, Matej Zecevic, Ilia Kuznetsov, Beatrix Suess 等ACL 2025
它引用的顶会 Paper4
- Big Bird: Transformers for Longer SequencesManzil Zaheer, Guru Guruganesh, Kumar Avinava Dubey, Joshua Ainslie 等NeurIPS 2020 · 被引用 3,159 次
- Multi-Stage Pre-training for Automated Chinese Essay ScoringWei Song, Kai Zhang, Ruiji Fu, Lizhen Liu 等EMNLP 2020 · 被引用 28 次
- PMAES: Prompt-mapping Contrastive Learning for Cross-prompt Automated Essay ScoringYuan Chen, Xia LiACL 2023 · 被引用 20 次
- TCFLE-8: a Corpus of Learner Written Productions for French as a Foreign Language and its Application to Automated Essay ScoringRodrigo Wilkens, Alice Pintard, David Alfter, Vincent Folny 等EMNLP 2023
相关 Paper
- Automated Cross-prompt Scoring of Essay TraitsRobert Ridley, Liang He, Xin-Yu Dai, Shujian Huang 等AAAI 2021 · 被引用 101 次
- Improving Domain Generalization for Prompt-Aware Essay Scoring via Disentangled Representation LearningZhiwei Jiang, Tianyi Gao, Yafeng Yin, Meng Liu 等ACL 2023 · 被引用 16 次
- Mixture of Ordered Scoring Experts for Cross-prompt Essay Trait ScoringPo-Kai Chen, Bo-Wei Tsai, Shao-Kuan Wei, Chien-Yao Wang 等ACL 2025
- Domain-Adaptive Neural Automated Essay ScoringYue Cao, Hanqi Jin, Xiaojun Wan, Zhiwei YuSIGIR 2020 · 被引用 47 次
- Autoregressive Multi-trait Essay Scoring via Reinforcement Learning with Scoring-aware Multiple RewardsHeejin Do, Sangwon Ryu, Gary Geunbae LeeEMNLP 2024 · 被引用 5 次
