CASEbot: A Conversational Agent for Structuring and Personalizing the Design of Self-Experiments in Personal Health
Sabrina Zaman Ishita, Sidharth Kaliappan, Mashrur Rashik, Daniel A. Epstein, Ravi Karkar
Abstract
Self-experimentation, or using tracked data to systematically answer health and wellbeing questions via hypothesis testing, has significant potential to support personal health. However, technological support for self-experimentation has focused on expert-designed self-experiments for specific health conditions, limiting people's ability to design their own rigorous experiments. To address this gap, we developed CASEbot (Conversation Agent for Self-Experimentation), an LLM-powered chatbot using a theory-driven approach to guide users through designing well-structured, personalized, and safe self-experiments. We conducted a within-subjects, mixed-methods study with 42 participants comparing CASEbot to a traditional worksheet-based approach. When formally comparing the experiment rigor and specificity, most participants designed better experiments using CASEbot. They appreciated CASEbot's conversational approach, which prompted them to surface everyday constraints and proactively raised safety concerns, but some found the platform too rigid in its recommendations. We discuss opportunities for future generative AI self-experimentation systems for health to balance structured guidance with user autonomy.
• Human-centered computing → Interactive systems and tools; User studies; Empirical studies in HCI.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 5b782b04-46b7-4e11-9284-8b793063f75eBuilds on24
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma et al.NeurIPS 2022 · 22,562 citations
- Retrieval-Augmented Generation for Knowledge-Intensive NLP TasksPatrick Lewis, Ethan Perez, Aleksandra Piktus, Fabio Petroni et al.NeurIPS 2020 · 19,162 citations
- MediQ: Question-Asking LLMs and a Benchmark for Reliable Interactive Clinical ReasoningShuyue Stella Li, Vidhisha Balachandran, Shangbin Feng, Jonathan Ilgen et al.NeurIPS 2024 · 215 citations
- Data-efficient Fine-tuning for LLM-based RecommendationXinyu Lin, Wenjie Wang, Yongqi Li, Shuo Yang et al.SIGIR 2024 · 152 citations
- Revisiting Reflection in HCI: Four Design Resources for Technologies that Support ReflectionMarit Bentvelzen, Pawel W. Wozniak, Pia S. F. Herbes, Evropi Stefanidi et al.UbiComp 2022 · 135 citations
Related papers
- Large Language Model Agents for Improving Engagement with Behavior Change Interventions: Application to Digital MindfulnessHarsh Kumar, Suhyeon Yoo, Angela M. Zavaleta Bernuy, Jiakai Shi et al.CSCW 2025 · 7 citations
- Customizing Emotional Support: How Do Individuals Construct and Interact With LLM-Powered ChatbotsXi Zheng, Zhuoyang Li, Xinning Gui, Yuhan LuoCHI 2025 · 47 citations
- Malicious LLM-Based Conversational AI Makes Users Reveal Personal InformationXiao Zhan, Juan Carlos Carrillo, William Seymour, Jose SuchUSENIX Security 2025
- Engagements with Generative AI and Personal Health Informatics: Opportunities for Planning, Tracking, Reflecting, and Acting around Personal Health DataShaan Chopra, Katherine Juarez, James Fogarty, Sean A. MunsonUbiComp 2025 · 11 citations
- Does My Chatbot Have an Agenda? Understanding Human and AI Agency in Human-Human-like Chatbot InteractionBhada Yun, Evgenia Taranova, April Yi WangCHI 2026 · 3 citations
