"Bespoke Bots": Diverse Instructor Needs for Customizing Generative AI Classroom Chatbots
Irene Hou, Zeyu Xiong, Philip J. Guo, April Yi Wang
Abstract
Instructors are increasingly experimenting with AI chatbots for classroom support. To investigate how instructors adapt chatbots to their own contexts, we first analyzed existing resources that provide prompts for educational purposes. We identified ten common categories of customization, such as persona, guardrails, and personalization. We then conducted interviews with ten university STEM instructors and asked them to card-sort the categories into priorities. We found that instructors consistently prioritized the ability to customize chatbot behavior to align with course materials and pedagogical strategies and de-prioritized customizing persona/tone. However, their prioritization of other categories varied significantly by course size, discipline, and teaching style, even across courses taught by the same individual, highlighting that no single design can meet all contexts. These findings suggest that modular AI chatbots may provide a promising path forward. We offer design implications for educational developers building the next generation of customizable classroom AI systems.
• Human-centered computing → Empirical studies in HCI.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b93b0d75-be86-4451-8b93-6ce0d850b943Builds on6
- Retrieval-Augmented Generation for Knowledge-Intensive NLP TasksPatrick Lewis, Ethan Perez, Aleksandra Piktus, Fabio Petroni et al.NeurIPS 2020 · 19,162 citations
- CodeAid: Evaluating a Classroom Deployment of an LLM-based Programming Assistant that Balances Student and Educator NeedsMajeed Kazemitabaar, Runlong Ye, Xiaoning Wang, Austin Zachary Henley et al.CHI 2024 · 246 citations
- ChainForge: A Visual Toolkit for Prompt Engineering and LLM Hypothesis TestingIan Arawjo, Chelse Swoopes, Priyan Vaithilingam, Martin Wattenberg et al.CHI 2024 · 141 citations
- EvalLM: Interactive Evaluation of Large Language Model Prompts on User-Defined CriteriaTae Soo Kim, Yoonjoo Lee, Jamin Shin, Young-Ho Kim et al.CHI 2024 · 81 citations
- A Piece of Theatre: Investigating How Teachers Design LLM Chatbots to Assist Adolescent Cyberbullying EducationMichael A. Hedderich, Natalie N. Bazarova, Wenting Zou, Ryun Shim et al.CHI 2024 · 42 citations
Related papers
- Exploring the Impact of Avatar Representations in AI Chatbot Tutors on Learning ExperiencesChek Tien Tan, Indriyati Atmosukarto, Budianto Tandianus, Songjia Shen et al.CHI 2025 · 16 citations
- Barriers that Programming Instructors Face While Performing Emergency Pedagogical Design to Shape Student-AI Interactions with Generative AI ToolsSam Lau, Kianoosh Boroojeni, Harry Keeling, Jenn MarroquinCHI 2026 · 1 citation
- CloChat: Understanding How People Customize, Interact, and Experience Personas in Large Language ModelsJuhye Ha, Hyeon Jeon, DaEun Han, Jinwook Seo et al.CHI 2024 · 66 citations
- Toward Scalable and Responsible Integration of Course-Specific AI Tutors: Instructor Experiences with a Campus-Wide PlatformEunhye Grace Ko, Hakeoung Hannah Lee, Anjali Singh, Lily Boddy et al.CHI 2026 · 2 citations
- Campus AI vs. Commercial AI: Comparing How Students and Employees Perceive their University's LLM Chatbot vs. ChatGPTLeon Hannig, Annika Bush, Meltem Aksoy, Tim Trappen et al.CHI 2026
