Evaluating Non-AI Experts' Interaction with AI: A Case Study In Library Context
Qingxiao Zheng, Minrui Chen, Hyanghee Park, Zhongwei Xu, Yun Huang
摘要
Public libraries in the U.S. are increasingly facing labor shortages, tight budgets, and overworked staff, creating a pressing need for conversational agents to assist patrons. The democratization of generative AI has empowered public service professionals to develop AI agents by leveraging large language models. To understand the needs of non-AI library professionals in creating their own conversational agents, we conducted semi-structured interviews with library professionals (n=11) across the U.S. Insights from these interviews informed the design of AgentBuilder, a prototype tool that enables non-AI experts to create conversational agents without coding skills. We then conducted think-aloud sessions and follow-up interviews to evaluate the prototype experience and identify the key evaluation criteria emphasized by library professionals (n=12) when developing conversational agents. Our findings highlight how these professionals perceive the prototype experience and reveal five essential evaluation criteria: interpreting user intent, faithful paraphrasing, proper alignment with authoritative sources, tailoring the tone of voice, and handling unknown answers effectively. These insights provide valuable guidance for designing AI-supported "end-user creation tools" in public service domains beyond libraries.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- Why Johnny Can't Prompt: How Non-AI Experts Try (and Fail) to Design LLM PromptsJ. D. Zamfirescu-Pereira, Richmond Y. Wong, Bjoern Hartmann, Qian YangCHI 2023 · 被引用 892 次
- Offline Training of Language Model Agents with Functions as Learnable WeightsShaokun Zhang, Jieyu Zhang, Jiale Liu, Linxin Song 等ICML 2024 · 被引用 41 次
- CloChat: Understanding How People Customize, Interact, and Experience Personas in Large Language ModelsJuhye Ha, Hyeon Jeon, DaEun Han, Jinwook Seo 等CHI 2024 · 被引用 66 次
- If I Hear You Correctly: Building and Evaluating Interview Chatbots with Active Listening SkillsZiang Xiao, Michelle X. Zhou, Wenxi Chen, Huahai Yang 等CHI 2020 · 被引用 113 次
- Pub-LawBench: Public-Oriented Benchmarking for LegalAIQiaoyu Zheng, Zehan Ma, Yijing Zhang, Qiqi Wang 等ACL 2026
