TCFLE-8: a Corpus of Learner Written Productions for French as a Foreign Language and its Application to Automated Essay Scoring
Rodrigo Wilkens, Alice Pintard, David Alfter, Vincent Folny, Thomas François
摘要
Automated Essay Scoring (AES) aims to automatically assess the quality of essays. Automation enables large-scale assessment, improvements in consistency, reliability, and standardization. Those characteristics are of particular relevance in the context of language certification exams. However, a major bottleneck in the development of AES systems is the availability of corpora, which, unfortunately, are scarce, especially for languages other than English. In this paper, we aim to foster the development of AES for French by providing the TCFLE-8 corpus, a corpus of 6.5k essays collected in the context of the Test de Connaissance du Français (TCF -French Knowledge Test) certification exam. We report the strict quality procedure that led to the scoring of each essay by at least two raters according to the levels of the Common European Framework of Reference for Languages (CEFR) and to the creation of a balanced corpus. In addition, we describe how linguistic properties of the essays relate to the learners' proficiency in TCFLE-8. We also advance the state-of-the-art performance for the AES task in French by experimenting with two strong baselines (i.e., RoBERTa and featurebased). Finally, we discuss the challenges of AES using TCFLE-8. 1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- DREsS: Dataset for Rubric-based Essay Scoring on EFL WritingHaneul Yoo, Jieun Han, So-Yeon Ahn, Alice OhACL 2025 · 被引用 13 次
- UniversalCEFR: Enabling Open Multilingual Research on Language Proficiency AssessmentJoseph Marvin Imperial, Abdullah Barayan, Regina Stodden, Rodrigo Wilkens 等EMNLP 2025 · 被引用 2 次
它引用的顶会 Paper1
相关 Paper
- Conundrums in Cross-Prompt Automated Essay Scoring: Making Sense of the State of the ArtShengjie Li, Vincent NgACL 2024 · 被引用 8 次
- Cross-Prompt Automated Essay Scoring of Multiple Traits: Making Sense of the State of the ArtShengjie Li, Vincent NgACL 2026 · 被引用 10 次
- CEFR-Based Sentence Difficulty Annotation and AssessmentYuki Arase, Satoru Uchida, Tomoyuki KajiwaraEMNLP 2022 · 被引用 17 次
- Improving Domain Generalization for Prompt-Aware Essay Scoring via Disentangled Representation LearningZhiwei Jiang, Tianyi Gao, Yafeng Yin, Meng Liu 等ACL 2023 · 被引用 16 次
- Mixture of Ordered Scoring Experts for Cross-prompt Essay Trait ScoringPo-Kai Chen, Bo-Wei Tsai, Shao-Kuan Wei, Chien-Yao Wang 等ACL 2025
