TeachTune: Reviewing Pedagogical Agents Against Diverse Student Profiles with Simulated Students
Hyoungwook Jin, Minju Yoo, Jeongeon Park, Yokyung Lee, Xu Wang, Juho Kim
摘要
Large language models (LLMs) can empower teachers to build pedagogical conversational agents (PCAs) customized for their students. As students have different prior knowledge and motivation levels, teachers must review the adaptivity of their PCAs to diverse students. Existing chatbot reviewing methods (e.g., direct chat and benchmarks) are either manually intensive for multiple iterations or limited to testing only single-turn interactions. We present TeachTune, where teachers can create simulated students and review PCAs by observing automated chats between PCAs and simulated students. Our technical pipeline instructs an LLM-based student to simulate prescribed knowledge levels and traits, helping teachers explore diverse conversation patterns. Our pipeline could produce simulated students whose behaviors correlate highly to their input knowledge and motivation levels within 5% and 10% accuracy gaps. Thirty science teachers designed PCAs in a between-subjects study, and using TeachTune resulted in a lower task load and higher student profile coverage over a baseline.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- EducationQ: Evaluating LLMs' Teaching Capabilities Through Multi-Agent Dialogue FrameworkYao Shi, Rongkeng Liang, Yong XuACL 2025 · 被引用 18 次
- From EduVisBench to EduVisAgent: A Benchmark and Multi-Agent Framework for Reasoning-Driven Pedagogical VisualizationHaonian Ji, Shi Qiu, Siyang Xin, Siwei Han 等ICLR 2026 · 被引用 6 次
- PLAID: Supporting Computing Instructors to Identify Domain-Specific Programming Plans at ScaleYoshee Jain, Mehmet Arif Demirtas, Kathryn I. CunninghamCHI 2025 · 被引用 2 次
- Enhancing Response Quality by Children in Voice-based Sleep Diaries via AI-based Continuous FeedbackShanshan Chen, Jun Hu, Gubing Wang, Panos MarkopoulosCHI 2026 · 被引用 2 次
- Exploring the Design and Impact of Interactive Worked Examples for Learners with Varying Prior KnowledgeSutapa Dey Tithi, Xiaoyi Tian, Ally Limke, Min Chi 等CHI 2026 · 被引用 1 次
它引用的顶会 Paper25
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma 等NeurIPS 2022 · 被引用 22,562 次
- Tree of Thoughts: Deliberate Problem Solving with Large Language ModelsShunyu Yao, Dian Yu, Jeffrey Zhao, Izhak Shafran 等NeurIPS 2023 · 被引用 5,068 次
- Generative Agents: Interactive Simulacra of Human BehaviorJoon Sung Park, Joseph C. O'Brien, Carrie Jun Cai, Meredith Ringel Morris 等UIST 2023 · 被引用 1,882 次
- Why Johnny Can't Prompt: How Non-AI Experts Try (and Fail) to Design LLM PromptsJ. D. Zamfirescu-Pereira, Richmond Y. Wong, Bjoern Hartmann, Qian YangCHI 2023 · 被引用 892 次
相关 Paper
- Simulated Students in Tutoring Dialogues: Substance or Illusion?Alexander Scarlatos, Jaewook Lee, Simon Woodhead, Andrew LanACL 2026 · 被引用 6 次
- Personality-aware Student Simulation for Conversational Intelligent Tutoring SystemsZhengyuan Liu, Stella Xin Yin, Geyu Lin, Nancy F. ChenEMNLP 2024 · 被引用 13 次
- Teach AI How to Code: Using Large Language Models as Teachable Agents for Programming EducationHyoungwook Jin, Seonghee Lee, Hyungyu Shin, Juho KimCHI 2024 · 被引用 94 次
- Exploring Teacher-Chatbot Interaction and Affect in Block-Based ProgrammingBahare Riahi, Ally Limke, Xiaoyi Tian, Viktoriia Storozhevykh 等CHI 2026 · 被引用 2 次
- CloChat: Understanding How People Customize, Interact, and Experience Personas in Large Language ModelsJuhye Ha, Hyeon Jeon, DaEun Han, Jinwook Seo 等CHI 2024 · 被引用 66 次
