Prioritising Hospital Complaints: An Innovative Tool Using Large Language Model-Assisted Content Analysis and Machine Learning Algorithms
Sulaiman, M. H.; Muda, N.; Abdul Razak, F.
Show abstract
BackgroundIn clinical settings, patients often express dissatisfaction through narrative speech or written text. However, most complaints management systems still rely on manual review or rulebased methods that fail to capture the severity or urgency of complaints. This leads to inconsistent triage, delayed resolution and missed opportunities for systemic improvement. A novel model leveraging large language model-assisted content analysis (LACA) and machine learning (ML) can transform subjective narratives into standardized, machine-readable severity scores, facilitating the prioritisation of complaints. ObjectiveThis study aims to (1) determine the precision, recall fscore and accuracy of the proposed predictive models used to classify comments into low-alert and high-alert comment, (2) determine the construct validity and internal consistency (Cronbachs ) of the themes found in LACA conducted on hospital web-based review data, (3) determine the predictors of low-alert and high-alert comments and their ability to change the log-odds of the outcome in logistic regression, and (4) to measure the robustness of the explanatory model measured by pseudo-R2. MethodologyLACA was performed using a set of thematic codes to generate an independent variable dataset (x), with a scale of 0: not an issue, 1: a small issue, 2: a moderate issue, 3: a serious issue, and 4: an extremely serious issue. The independent variables (x) and the dependent variable (y, representing the review rating) were then split into training and testing sets to build predictive ML models. Grid search was used to determine the optimal combination of hyperparameters. The performance of the predictive and explanatory models was evaluated. ResultsML classification was able to produce f1-score of 0.88 - 0.94 and accuracy of 0.92 for LR model; and f1-score of 0.87 - 0.94 and accuracy of 0.92 for ANN model The behaviour of predictive models was successfully explained by the explanatory model: Six (6) themes were determined with cumulative explained variance (CEV) of 0.74 and average Cronbachs of 0.86. LR shows significance on 5 themes with pseudo-R2 of 0.55. ConclusionThis study demonstrates that a data pipeline utilizing LACA and ML algorithms shows excellent performance in classifying patient comments in a hospital setting. All effectiveness parameters including CEV, Cronbachs , precision, recall, f1-score, and accuracy indicate strong performance in differentiating high-alert from low-alert comments.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Impact of electronic medical records on healthcare delivery in Nigeria: A Review 94%
- Identification of predictive patient characteristics for assessing the probability of COVID-19 in-hospital mortality 94%
- Comparing human vs. machine-assisted analysis to develop a new approach for Big Qualitative Data Analysis 94%
Similar papers in this journal
- Common misconceptions held by health researchers when interpreting linear regression assumptions, a cross-sectional study 94%
- Accuracy of online symptom checkers and the potential impact on service utilisation 94%
- Essential Indicators of Quality in Primary Care Settings: An Evidence-Based, Structured, Expert Approach 93%
Similar papers in this journal
- The role of IT ambidexterity, digital dynamic capability and knowledge processes as enablers of patient agility: an empirical study 94%
- Learning from the resilience of hospitals and their staff to the COVID-19 pandemic: a scoping review 93%
- Prediction of COVID-19 Mortality to Support Patient Prognosis and Triage and Limits of Current Open-Source Data 92%
Similar papers in this journal
- Health indicators as a measure of individual health status: public perspectives 94%
- Artificial Intelligence (AI)-based Chatbots in Promoting Health Behavioral Changes: A Systematic Review 93%
- Structured Codes and Free-Text Notes: Measuring Information Complementarity in Electronic Health Records 93%
Similar papers in this journal
- The potential for digital patient symptom recording through symptom assessment applications to optimize patient flow and reduce waiting times in Urgent Care Centers: a simulation study 95%
- A Web-based, Mobile Responsive Application to Screen Healthcare Workers for COVID Symptoms: Descriptive Study 94%
- Exploring Patient and Staff Experiences of Video Consultations During COVID-19 in an English Outpatient Care Setting: Secondary Data Analysis of Routinely Collected Feedback Data 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.