Thematic-LM: A LLM-based Multi-agent System for Large-scale Thematic Analysis
Tingrui Qiao, Caroline Walker, Chris Cunningham, Yun Sing Koh
摘要
Thematic analysis (TA) is a widely used qualitative method for identifying underlying meanings within unstructured text. However, TA requires manual processes, which become increasingly labour-intensive and time-consuming as datasets grow. While large language models (LLMs) have been introduced to assist with TA on small-scale datasets, three key limitations hinder their effectiveness. First, current approaches often depend on interactions between an LLM agent and a human coder, a process that becomes challenging with larger datasets. Second, with feedback from the human coder, the LLM tends to mirror the human coder, which provides a narrower viewpoint of the data. Third, existing methods follow a sequential process, where codes are generated for individual samples without recalling previous codes and associated data, reducing the ability to analyse data holistically. To address these limitations, we propose Thematic-LM, an LLM-based multi-agent system for large-scale computational thematic analysis. Thematic-LM assigns specialised tasks to each agent, such as coding, aggregating codes, and maintaining and updating the codebook. We assign coder agents different identity perspectives to simulate the subjective nature of TA, fostering a more diverse interpretation of the data. We applied Thematic-LM to the Dreaddit dataset and the Reddit climate change dataset to analyse themes related to social media stress and online opinions on climate change. We evaluate the resulting themes based on trustworthiness principles in qualitative research. Our study reveals insights such as assigning different identities to coder agents promotes divergence in codes and themes.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper3
- Many Minds, One Goal: Time Series Forecasting via Sub-task Specialization and Inter-agent CooperationQihe Huang, Zhengyang Zhou, Yangze Li, Kuo Yang 等NeurIPS 2025 · 被引用 11 次
- Personalized Federated Fine-Tuning for LLMs via Data-Driven Heterogeneous Model ArchitecturesYicheng Zhang, Zhen Qin, Zhaomin Wu, Jian Hou 等WWW 2026 · 被引用 9 次
- What Users Ask, Policies Miss: Unveiling the Gap Between Community-Expressed Privacy Concerns and LLM Provider PoliciesZhihuang Liu, Zhen Huang, Ling Hu, Yifan Yang 等USENIX Security 2026
相关 Paper
- LATA: A Pilot Study on LLM-Assisted Thematic Analysis of Online Social Network Data Generation ExperiencesQile Wang, Moath Erqsous, Kenneth E. Barner, Matthew Louis MaurielloCSCW 2025 · 被引用 20 次
- SCALE: Towards Collaborative Content Analysis in Social Science with Large Language Model Agents and Human InterventionChengshuai Zhao, Zhen Tan, Chau-Wai Wong, Xinyan Zhao 等ACL 2025 · 被引用 8 次
- ThemeViz: Understanding the Effect of Human-AI Collaboration in Theme Development with an LLM-enhanced Interactive Visual SystemDaye Kang, Zhuolun Han, Jiahe Tian, Muhan Zhang 等CSCW 2025 · 被引用 2 次
- Reimagining Support: Exploring Autistic Individuals' Visions for AI in Coping with Negative Self-TalkBuse Çarik, Victoria V. Izaac, Xiaohan Ding, Angela Scarpa 等CHI 2025 · 被引用 19 次
- Cinema Multiverse Lounge: Enhancing Film Appreciation via Multi-Agent ConversationsJeongwoo Ryu, Kyusik Kim, Dongseok Heo, Hyungwoo Song 等CHI 2025 · 被引用 9 次
