Into the Unknown Unknowns: Engaged Human Learning through Participation in Language Model Agent Conversations
Yucheng Jiang, Yijia Shao, Dekun Ma, Sina J. Semnani, Monica S. Lam
Abstract
While language model (LM)-powered chatbots and generative search engines excel at answering concrete queries, discovering information in the terrain of unknown unknowns remains challenging for users. To emulate the common educational scenario where children/students learn by listening to and participating in conversations with their parents/teachers, we create Collaborative STORM (Co-STORM). 1 Unlike QA systems that require users to ask all the questions, Co-STORM lets users observe and occasionally steer the discourse among several LM agents. The agents ask questions on the user's behalf, allowing the user to discover unknown unknowns serendipitously. To facilitate user interaction, Co-STORM assists users in tracking the discourse by organizing the uncovered information into a dynamic mind map, ultimately generating a comprehensive report as takeaways. For automatic evaluation, we construct the WildSeek dataset by collecting real information-seeking records with user goals. Co-STORM outperforms baseline methods on both discourse trace and report quality. In a further human evaluation 2 , 70% of participants prefer Co-STORM over a search engine, and 78% favor it over a RAG (Retrieval Augmented Generation) chatbot. * Equal Contribution 1 Our resources and code are released at https://github .com/stanford-oval/storm .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ed980a2f-cbb3-42d0-a1d3-4f93703025d3Cited by top-tier papers16
- MEM1: Learning to Synergize Memory and Reasoning for Efficient Long-Horizon AgentsZijian Zhou, Ao Qu, Zhaoxuan Wu, Sunghwan Kim et al.ICLR 2026 · 223 citations
- Beyond Outlining: Heterogeneous Recursive Planning for Adaptive Long-form Writing with Language ModelsRuibin Xiong, Yimeng Chen, Dmitrii Khizbullin, Mingchen Zhuge et al.EMNLP 2025 · 14 citations
- WikiAutoGen: Towards Multi-Modal Wikipedia-Style Article GenerationZhongyu Yang, Jun Chen, Dannong Xu, Junjie Fei et al.ICCV 2025 · 3 citations
- XtraGPT: Context-Aware and Controllable Academic Paper Revision via Human-AI CollaborationNuo Chen, Andre Huikai Lin, Jiaying Wu, Junyi Hou et al.ACL 2026 · 3 citations
- Writing Like the Best: Exemplar-Based Expository Text GenerationYuxiang Liu, Kevin Chen-Chuan ChangACL 2025 · 2 citations
Builds on12
- Retrieval-Augmented Generation for Knowledge-Intensive NLP TasksPatrick Lewis, Ethan Perez, Aleksandra Piktus, Fabio Petroni et al.NeurIPS 2020 · 19,162 citations
- CAMEL: Communicative Agents for "Mind" Exploration of Large Language Model SocietyGuohao Li, Hasan Hammoud, Hani Itani, Dmitrii Khizbullin et al.NeurIPS 2023 · 1,975 citations
- Generative Agents: Interactive Simulacra of Human BehaviorJoon Sung Park, Joseph C. O'Brien, Carrie Jun Cai, Meredith Ringel Morris et al.UIST 2023 · 1,882 citations
- RAPTOR: Recursive Abstractive Processing for Tree-Organized RetrievalParth Sarthi, Salman Abdullah, Aditi Tuli, Shubh Khanna et al.ICLR 2024 · 460 citations
- Encouraging Divergent Thinking in Large Language Models through Multi-Agent DebateTian Liang, Zhiwei He, Wenxiang Jiao, Xing Wang et al.EMNLP 2024 · 177 citations
Related papers
- How AI Processing Delays Foster Creativity: Exploring Research Question Co-Creation with an LLM-based AgentYiren Liu, Si Chen, Haocong Cheng, Mengxia Yu et al.CHI 2024 · 57 citations
- Ask and Retrieve Knowledge: Towards Proactive Asking with Imperfect Information in Medical Multi-turn DialoguesBolin Zhang, Shengwei Wang, Yangqin Jiang, Dianbo Sui et al.SIGIR 2025
- Shoot First, Ask Questions Later? Building Rational Agents that Explore and Act Like PeopleGabriel Grand, Valerio Pepe, Joshua B. Tenenbaum, Jacob AndreasICLR 2026 · 7 citations
- Feedback-Aware MCTS for Goal-Oriented Information SeekingHarshita Chopra, Chirag ShahNeurIPS 2025 · 2 citations
- MediQ: Question-Asking LLMs and a Benchmark for Reliable Interactive Clinical ReasoningShuyue Stella Li, Vidhisha Balachandran, Shangbin Feng, Jonathan Ilgen et al.NeurIPS 2024 · 215 citations
