CoGrader: Transforming Instructors' Assessment of Project Reports through Collaborative LLM Integration
Zixin Chen, Jiachen Wang, Yumeng Li, Haobo Li, Chuhan Shi, Rong Zhang, Huamin Qu
摘要
Grading project reports is increasingly significant in today's educational landscape, where they serve as key assessments of students' comprehensive problem-solving abilities. However, it remains challenging due to the multifaceted evaluation criteria involved, such as creativity and peer-comparative achievement. Meanwhile, instructors often struggle to maintain fairness throughout the timeconsuming grading process. Recent advances in AI, particularly large language models, have demonstrated potential for automating simpler grading tasks, such as assessing quizzes or basic writing quality. However, these tools often fall short when it comes to complex metrics, like design innovation and the practical application of knowledge, that require an instructor's educational insights and contextual understanding of the class. To address this challenge, we conducted a formative study with six instructors and developed CoGrader, which introduces a novel grading workflow combining human-LLM collaborative metrics design, benchmarking, and AI-assisted feedback. CoGrader was found effective in improving grading efficiency and consistency while providing reliable peercomparative feedback to students. We also discuss design insights and ethical considerations for the development of human-AI collaborative grading systems.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper21
- LLM Evaluators Recognize and Favor Their Own GenerationsArjun Panickssery, Samuel R. Bowman, Shi FengNeurIPS 2024 · 被引用 865 次
- The Impact of Generative AI on Critical Thinking: Self-Reported Reductions in Cognitive Effort and Confidence Effects From a Survey of Knowledge WorkersHao-Ping (Hank) Lee, Advait Sarkar, Lev Tankelevitch, Ian Drosos 等CHI 2025 · 被引用 690 次
- Benchmarking Large Language Models in Retrieval-Augmented GenerationJiawei Chen, Hongyu Lin, Xianpei Han, Le SunAAAI 2024 · 被引用 531 次
- CodeAid: Evaluating a Classroom Deployment of an LLM-based Programming Assistant that Balances Student and Educator NeedsMajeed Kazemitabaar, Runlong Ye, Xiaoning Wang, Austin Zachary Henley 等CHI 2024 · 被引用 246 次
- Human Creativity in the Age of LLMs: Randomized Experiments on Divergent and Convergent ThinkingHarsh Kumar, Jonathan Vincentius, Ewan Jordan, Ashton AndersonCHI 2025 · 被引用 107 次
相关 Paper
- Co-designing Large Language Model Tools for Project-Based Learning with K12 EducatorsPrerna Ravi, John Masla, Gisella Kakoti, Grace C. Lin 等CHI 2025 · 被引用 26 次
- Open-ended Structured Question Assessment with Human-LLM CollaborationFengyan Lin, Yanna Lin, Kai Cao, Zikun Deng 等CHI 2026 · 被引用 1 次
- From Replication to Redesign: Exploring Pairwise Comparisons for LLM-Based Peer ReviewYaohui Zhang, Haijing Zhang, Wenlong Ji, Tianyu Hua 等NeurIPS 2025 · 被引用 15 次
- CogBench: a large language model walks into a psychology labJulian Coda-Forno, Marcel Binz, Jane X. Wang, Eric SchulzICML 2024 · 被引用 60 次
- Charting the Future of AI in Project-Based Learning: A Co-Design Exploration with StudentsChengbo Zheng, Kangyu Yuan, Bingcan Guo, Reza Hadi Mogavi 等CHI 2024 · 被引用 65 次
