Using AI to support rapid qualitative data analysis of survey and interview data in public health: a proof-of-concept study.
White, T.; Pae, R.; Hoerst, C.; Stuart, A.; Abbey, R.; Skryabina, E.; Brooks, S. K.; Bondaronek, P.; Cattarino, L.
Show abstract
Background: Large language models (LLMs) are increasingly used to support a growing range of analytical and operational tasks in public health, and show promise for assisting qualitative text analysis at scale. While they have demonstrated utility in structured natural language processing tasks, their role in more interpretive approaches such as thematic analysis remains less clear. Thematic analysis requires careful, often time-intensive engagement with qualitative data, which can be challenging when datasets are large or when policy teams must deliver rapid insights. This study examines whether a pragmatic LLM-assisted workflow can support early-stage thematic analysis of large public health datasets in a way that is systematic, transparent, and compatible with human-led analytical oversight. Methods: We developed and evaluated a proof-of-concept workflow for LLM-assisted semantic coding and theme identification in a public health case study involving qualitative survey and debrief data. The LLM generated codes and themes using a consistent prompting structure. We compared those themes with those produced through human-only thematic analysis and with outputs from a topic-modelling based approach. A convergence coding matrix was used to classify alignment as agreement, complementarity, disagreement or silence. To evaluate the reliability of the LLM's theme classification, we manually labelled 500 codes and calculated accuracy metrics. Results: The LLM-generated themes showed broad alignment with the human analysis at the level of higher-order themes, with agreement observed for 73% of manual themes. Complementary LLM themes were found for a further 18%, while only one subtheme showed dissonance and a small number (8.3%) had no clear match. Compared with the topic modelling approach, the LLM produced a broader and more detailed thematic structure. In the theme assignment task, the LLM achieved an overall F1 score of 0.68, with individual theme scores between 0.34 and 0.80, indicating moderate consistency with humans and stronger performance for clearly defined themes. Conclusions: These findings suggest that LLM-assisted thematic analysis has promise as a pragmatic proof-of-concept approach for rapid, higher-level qualitative sensemaking in public health, particularly when datasets are large and timely insight is needed. Its use should remain bounded to surface-level exploratory analysis and embedded within human-led workflows.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Comparing human vs. machine-assisted analysis to develop a new approach for Big Qualitative Data Analysis 94%
- From Patient Voices to Policy: Data Analytics Reveals Patterns in Ontarios Hospital Feedback 92%
- Conversational, Longitudinal, Ecological Assessment (CLEA): Exploring a new AI-driven method for qualitative data collection in a behavioural health context 92%
Similar papers in this journal
- Wellbeing Impact Study of High-Speed 2 (WISH2): Protocol for a mixed-methods examination of the impact of major transport infrastructure development on mental health and wellbeing 93%
- Stakeholders’ views on an institutional dashboard with metrics for responsible research 92%
- Knowledge mobilisation of rapid evidence reviews to inform health and social care policy and practice in a public health emergency: appraisal of the Wales COVID-19 Evidence Centre processes and impact, 2021-23 92%
Similar papers in this journal
- Subnational tailoring of malaria interventions for strategic planning and prioritization: experience and perspectives of five malaria programs 92%
- How does policy modelling work in practice? A global analysis on the use of epidemiological modelling in health crises 92%
- Advancing the Safe Motherhood Initiative: a qualitative and sentiment analysis of local physician’s perspectives on antibiotic self-medication during pregnancy in a low- and middle-income country 91%
Similar papers in this journal
- Ethnicity and COVID-19 outcomes among healthcare workers in the United Kingdom: UK-REACH ethico-legal research, qualitative research on healthcare workers’ experiences, and stakeholder engagement protocol 94%
- Evaluating a first fully automated interview grounded in Multiple Mini Interview (MMI) methodology: results from a feasibility study 92%
- Mapping factors that may influence attrition and retention of midwives: a scoping review protocol 91%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.