"Don't Forget the Teachers": Towards an Educator-Centered Understanding of Harms from Large Language Models in Education
Emma Harvey, Allison Koenecke, René F. Kizilcec
Abstract
Education technologies (edtech) are increasingly incorporating new features built on large language models (LLMs), with the goals of enriching the processes of teaching and learning and ultimately improving learning outcomes. However, the potential downstream impacts of LLM-based edtech remain understudied. Prior attempts to map the risks of LLMs have not been tailored to education specifically, even though it is a unique domain in many respects: from its population (students are often children, who can be especially impacted by technology) to its goals (providing the correct answer may be less important for learners than understanding how to arrive at an answer) to its implications for higher-order skills that generalize across contexts (e.g., critical thinking and collaboration). We conducted semi-structured interviews with six edtech providers representing leaders in the K-12 space, as well as a diverse group of 23 educators with varying levels of experience with LLM-based edtech. Through a thematic analysis, we explored how each group is anticipating, observing, and accounting for potential harms from LLMs in education. We find that, while edtech providers focus primarily on mitigating technical harms, i.e., those that can be measured based solely on LLM outputs themselves, educators are more concerned about harms that result from the broader impacts of LLMs, i.e., those that require observation of interactions between students, educators, school systems, and edtech to measure. Overall, we (1) develop an education-specific overview of potential harms from LLMs, (2) highlight gaps between conceptions of harm by edtech providers and those by educators, and (3) make recommendations to facilitate the centering of educators in the design and development of edtech tools.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers14
- Beyond In-Domain Detection: SpikeScore for Cross-Domain Hallucination DetectionYongxin Deng, Zhen Fang, Sharon Li, Ling ChenICLR 2026 · 5 citations
- Do Teachers Dream of GenAI Widening Educational (In)equality? Envisioning the Future of K-12 GenAI Education from Global Teachers' PerspectivesRuiwei Xiao, Qing Xiao, Xinying Hou, Phenyo Phemelo Moletsane et al.CHI 2026 · 4 citations
- An Empirical Study to Understand How Students Use ChatGPT for Writing EssaysAndrew Jelson, Daniel Manesh, Alice Jang, Daniel Dunlap et al.CHI 2026 · 3 citations
- Exploring Teacher-Chatbot Interaction and Affect in Block-Based ProgrammingBahare Riahi, Ally Limke, Xiaoyi Tian, Viktoriia Storozhevykh et al.CHI 2026 · 2 citations
- MusicScaffold: Bridging Machine Efficiency and Human Growth in Adolescent Creative Education through Generative AIZhejing Hu, Yan Liu, Zhi Zhang, Gong Chen et al.CHI 2026 · 2 citations
Builds on17
- Retrieval-Augmented Generation for Knowledge-Intensive NLP TasksPatrick Lewis, Ethan Perez, Aleksandra Piktus, Fabio Petroni et al.NeurIPS 2020 · 19,162 citations
- Teachers, Parents, and Students' perspectives on Integrating Generative AI into Elementary Literacy EducationAriel Han, Xiaofei Zhou, Zhenyao Cai, Shenshen Han et al.CHI 2024 · 104 citations
- Mathemyths: Leveraging Large Language Models to Teach Mathematical Language through Child-AI Co-Creative StorytellingChao Zhang, Xuechen Liu, Katherine Ziska, Soobin Jeon et al.CHI 2024 · 93 citations
- The Promise and Peril of ChatGPT in Higher Education: Opportunities, Challenges, and Design ImplicationsHyanghee Park, Daehwan AhnCHI 2024 · 80 citations
- The Situate AI Guidebook: Co-Designing a Toolkit to Support Multi-Stakeholder, Early-stage Deliberations Around Public Sector AI ProposalsAnna Kawakami, Amanda Coston, Haiyi Zhu, Hoda Heidari et al.CHI 2024 · 54 citations
Related papers
- Understanding the Effect of Risk Perception on the Acceptance and Use of Large Language Models Among University StudentsMichael T. Rücker, Carolin Büchting, Thomas KoschCSCW 2025 · 4 citations
- Appraising the Potential Uses and Harms of LLMs for Medical Systematic ReviewsHye Sun Yun, Iain James Marshall, Thomas A. Trikalinos, Byron C. WallaceEMNLP 2023 · 11 citations
- Farsight: Fostering Responsible AI Awareness During AI Application PrototypingZijie J. Wang, Chinmay Kulkarni, Lauren Wilcox, Michael Terry et al.CHI 2024 · 55 citations
- K-12EduBench: A Benchmark for Evaluating Large Language Models' Knowledge, Problem-Solving, and Educational Goal Cognition in K-12 EducationYuqing Ye, Xuan Zhou, Zhifu Chen, Dandan Li et al.AAAI 2026
- Knowledge without Wisdom: Measuring Misalignment between LLMs and Intended ImpactMichael Hardy, Yunsung KimACL 2026
