Development and Validation of VC-MAES and VC-SEPS: Deep Learning-Based Early Warning Systems for Hospitalized Patients
Kim, Y.; Hahn, S.; Kim, K. J.; Yang, E.; Kim, J.-h.; Lee, S.; Han, C. H.; Won, J.-Y.; Ahn, B. E.; Mun, Y.; Chung, K. S.; Sim, T.
Show abstract
The timely detection of ward deterioration--including unplanned intensive care unit (ICU) transfer, cardiac arrest, death, and sepsis--remains an unmet need. Although rule-based early warning scores and newer machine-learning models have been introduced, their clinical adoption is limited owing to challenges such as low predictive performance, excessive false alarms, and lack of generalizability. This study aimed to develop two deep-learning models for patients in general wards using a bidirectional long short-term memory neural network architecture: (i) the VitalCare-Major Adverse Event Score (VC-MAES), which predicts clinical deterioration events (CDEs)-- unplanned ICU transfer, cardiac arrest, or in-hospital death--within 6 h and (ii) the VitalCare-SEPsis Score (VC-SEPS), which predicts sepsis onset within 4 h. Additionally, we sought to externally validate the performance of the models in an independent cohort. This study was conducted in two sequential phases. First, the VC-MAES and VC-SEPS models were developed using a large retrospective cohort from Yonsei Severance Hospital, Seoul, Republic of Korea. Second, external validation was performed in a single-center cohort at National Health Insurance Service Ilsan Hospital (NHIS Ilsan Hospital), Ilsan, Republic of Korea. Both algorithms incorporated patient age, vital signs, laboratory results, and the Glasgow Coma Scale scores. VC-MAES performance was compared with the Modified Early Warning Score (MEWS) and National Early Warning Score (NEWS), whereas VC-SEPS performance was compared with the Sequential Organ Failure Assessment (SOFA), quick SOFA (qSOFA), and NEWS, using the area under the receiver operating characteristic curve (AUROC) as the primary performance metric. The derivation cohort comprised 357,009 adult general-ward admissions at Yonsei Severance Hospital (2013-2017), and the external validation cohort included 22,073 admissions at NHIS Ilsan Hospital (2017). In the external validation cohort, the VC-MAES predicted CDEs within 6 h with an AUROC of 0.918 (95% confidence interval [CI], 0.909-0.927), outperforming the MEWS (0.834; 95% CI, 0.820-0.849) and NEWS (0.883; 95% CI, 0.869- 0.896). VC-SEPS predicted sepsis onset within 4 h, with an AUROC of 0.941 (95% CI, 0.934-0.947), surpassing the SOFA (0.559; 95% CI, 0.546-0.571), qSOFA (0.687; 95% CI, 0.671-0.704), and NEWS (0.767; 95% CI, 0.748-0.785). Both models maintained AUROC values above 0.86 across all age and sex categories. The VC-MAES and VC-SEPS outperformed conventional early warning scores in predicting clinical deterioration events and sepsis. These models can enable earlier and more precise interventions, enhancing patient care and optimizing resource use, ultimately leading to better patient outcomes without overwhelming healthcare providers.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Development and Prospective Implementation of a Large Language Model based System for Early Sepsis Prediction 94%
- GLUCOSE: A Distributional Reinforcement Learning Model for Optimal Glucose Control After Cardiac Surgery 93%
- CT-based Rapid Triage of COVID-19 Patients: Risk Prediction and Progression Estimation of ICU Admission, Mechanical Ventilation, and Death of Hospitalized Patients 93%
Similar papers in this journal
- A comparison of machine learning models versus clinical evaluation for mortality prediction in patients with sepsis 97%
- Development of a Risk Prediction Model for Sepsis-Related Delirium Based on Multiple Machine Learning Approaches and an Online Calculator 97%
- SOFA score performs worse than age for predicting mortality in patients with COVID-19 95%
Similar papers in this journal
- Generalizability Challenges of Mortality Risk Prediction Models: A Retrospective Analysis on a Multi-center Database 94%
- Identification of predictive patient characteristics for assessing the probability of COVID-19 in-hospital mortality 94%
- Geographical validation of the Smart Triage Model by age group 93%
Similar papers in this journal
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.