Automated Healthcare Thematic Analysis using Multi-Agent Large Language Model: Algorithm Development and Evaluation
cs.HC, cs.AI
Submitted: 2025-12-18
Updated: 2026-08-25
Comments: 45 pages, 5 figures
Journal ref: Published at JMIR in 2026
Code: https://github.com/QidiXu96/CoTI-MultiAgent-Theme-Analysis
Project page: https://maartengr.github.io/BERTopic/getting_started/supervised/supervised.html
License: http://creativecommons.org/licenses/by-nc-nd/4.0/
The gist: Understanding patients experiences is essential for advancing patient-centered care.
Terminology
Abstract
Understanding patients experiences is essential for advancing patient-centered care. Qualitative thematic analysis is widely used to explore these experiences, however, the process remains labor-intensive, subjective, and difficult to scale. This study aimed to develop and evaluate Collaborative Theme Identification Agent (CoTI), a multi-agent large language model framework designed to support manual thematic analysis by rapidly generating supporting excerpts, initial codes, and themes. CoTI consists of three agents: Instructor, Thematizer, and CodebookGenerator. The Instructor refines instruction prompts, the Thematizer extracts supporting excerpts and generates initial codes for each transcript, and the CodebookGenerator groups similar codes across all transcripts into a codebook with themes. We evaluated CoTI primarily using 12 heart failure patient transcripts. CoTI-generated outputs were compared against the reference standard developed by senior investigators. To explore human-AI interaction in thematic analysis, we further implemented CoTI in a user-facing application. CoTI generated supporting excerpts, initial codes, and themes that were more similar to those of senior investigators than did the outputs of junior investigators, baseline natural language processing models, and other basic large language models. In an exploratory human-AI collaboration experiment, we found that the collaboration between CoTI and junior investigators provided only marginal gains compared to CoTI alone. A possible hypothesis was that junior investigators may over-rely on CoTI and limit their independent critical thinking. CoTI can improve the efficiency of thematic analysis by rapidly generating supporting excerpts, initial codes, and themes for human researchers review. These findings highlight CoTI potential as a useful tool for scalable qualitative research.
Related papers
- EduGage: A Multimodal Dataset and Benchmark for Sensor-Based Momentary Assessment of Engagement in Self-Guided Video Learning
- EvoDesign: Agentic Editable Diagram Creation via Design Expertise Evolution
- HAGI++: Head-Assisted Gaze Imputation and Generation
- Linking Behaviour and Perception to Evaluate Meaningful Human Control over Partially Automated Driving
- Review of Explainable Decision Support and Adaptive Human-Machine Interfaces for Automation Transparency in Maritime Autonomous Surface Ships
- Towards Cognitive Process-Aware Proactive Writing Support