Clinical Sentiment Analysis by Large Language Models Enhances the Prediction of Hepatorenal Syndrome in Decompensated Cirrhosis
Lai, M.; Fenton, C.; Rubin, J.; Huang, C.-Y.; Pletcher, M. J.; Lai, J. C.; Cullaro, G.; Ge, J.
Show abstract
Background and AimsHepatorenal syndrome - Acute Kidney Injury (HRS-AKI) is a severe complication of decompensated cirrhosis that is challenging to predict. Sentiment analysis, a computational process of identifying and categorizing opinions and judgment expressed in text, may enhance traditional prediction methodologies based on structured variables. Large language models (LLMs), such as generative pretrained transformers (GPTs), have demonstrated abilities to perform sentiment analyses on non-clinical texts. We sought to determine if GPT-performed sentiment analysis could improve upon predictions using clinical covariates alone in the prediction of HRS-AKI. MethodsAdult patients admitted to a single academic medical center with decompensated cirrhosis and AKI. We used a protected health information (PHI) compliant version of Microsoft Azure OpenAI GPT-4o to derive a sentiment score ranging from 0 to 1 for HRS-AKI, and conduct natural language processing (NLP) extraction of clinical terms associated with HRS-AKI in clinical notes. The area under the receiver operator curve (AUROC) was compared in logistic regression models incorporating structured variables (socio-demographics, MELD 3.0, hemodynamic parameters) with compared to without sentiment scores and NLP-extracted clinical terms. ResultsIn our cohort of 314 participants, higher sentiment score was associated with the diagnosis of HRS-AKI (OR 1.33 per 0.1, 95% CI 1.02-1.79) in multivariate models. AUROC of the baseline model using structured clinical covariates alone was 0.639. With the addition of the GPT-4o derived sentiment score and clinical terms to structured covariates, the final model yielded an improved AUROC of 0.758 (p= 0.03). ConclusionsClinical texts contain large amounts of data that are currently difficult to extract using standard methodologies. Sentiment analysis and NLP-based variable derivation with GPT-4o in clinical application is feasible and can improve the prediction of HRS-AKI over traditional modeling methodologies alone.
Matching journals
The top 8 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- AKI Risk Score (AKI-RiSc): Developing an Interpretable Clinical Score for Early Identification of Acute Kidney Injury for Patients Presenting to the Emergency Department 95%
- Developing And Validating COVID-19 Adverse Outcome Risk Prediction Models From A Bi-National European Cohort Of 5594 Patients 93%
- Application of physiological network mapping in the prediction of survival in critically ill patients with acute liver failure 93%
Similar papers in this journal
- Evaluating the kidney disease progression using a comprehensive patient profiling algorithm: A hybrid clustering approach 94%
- Predicting Short-Term Mortality in Severe Cirrhosis: An Interpretable Machine Learning Model Integrating Routine Clinical Indicators 94%
- A comparison of machine learning models versus clinical evaluation for mortality prediction in patients with sepsis 93%
Similar papers in this journal
- Identification of predictive patient characteristics for assessing the probability of COVID-19 in-hospital mortality 91%
- Predictability and Stability Testing to Assess Clinical Decision Instrument Performance for Children After Blunt Torso Trauma 91%
- Use of a Continuous Single Lead Electrocardiogram Analytic to Predict Patient Deterioration Requiring Rapid Response Team Activation 91%
Similar papers in this journal
- Clinical course and risk factors for mortality of COVID-19 patients with pre-existing cirrhosis: A multicenter cohort study 92%
- Optimising Early Management of Acute Severe Ulcerative Colitis in the Biologics Era: Admission Model for Intensification of Therapy in Acute Severe Colitis (ADMIT–ASC) 89%
- Sphincterotomy for Biliary Sphincter of Oddi Disorder and idiopathic Acute Recurrent Pancreatitis: THE RESPOND LONGITUDINAL COHORT 88%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.