Back

Outcome Prediction Models for Critically Ill Patients Using Small Routine Laboratory Datasets

Cao, X.; Hou, J.; Wei, X.; Wang, Q.

2026-04-27 emergency medicine
10.64898/2026.04.26.26351758 medRxiv
Show abstract

We present a suite of foundational, outcome prediction models for critically ill patients, developed using readily available, routine blood tests and advanced machine learning techniques. The input data of the models includes complete blood counts (CBCs), metabolic panels, and additional biomarkers that assess liver and kidney function, coagulation status, and cardiac injury. The output yields the predicted outcome at a given future horizon. For diagnoses, the length of the future horizon is set to zero while it is set to a fixed time interval for prognoses. The training dataset in this study comprises clinical data from 332 ICU patients, augmented with 200 synthetic samples generated via a conditional diffusion model. Generative machine learning-based data imputation and augmentation approaches yielded modest gains in predictive accuracy. However, substantial performance improvements were achieved through additional methods, including dimensionality and order reduction, SHAP-based feature importance analysis, and a novel time-series-to-image encoding strategy that enables the use of image-based classifiers for temporal clinical data. Principal component analysis-based order reduction produced measurable gains in outcome prediction, while the time-series-to-image encoding proved particularly effective in mitigating small-data limitations common in clinical research. Across all evaluation metrics--accuracy, precision, recall, F1 score, and AUROC--the prognostic models achieved performance exceeding 85%, with some models attaining AUROC scores above 90%. We innovated a new model-ensemble approach to optimize the predictive outcome. This ensemble modeling approach improves the overal prediction, pushing all assessment metrics over 90%. This work establishes a robust and interpretable AI-enabled diagnostic and prognostic toolkit for outcome predictions in critically ill patients and demonstrates a scalable workflow for developing high-performing models from sparse healthcare datasets. The proposed framework is readily deployable in ICU environments with routine blood testing capabilities and serves as a foundation for future integration into digital twin systems for critical care.

Matching journals

The top 3 journals account for 50% of the predicted probability mass.

1
Scientific Reports
3612 papers in training set
Top 0.6%
19.0%
2
PLOS Digital Health
106 papers in training set
Top 0.2%
19.0%
3
npj Digital Medicine
118 papers in training set
Top 0.5%
12.2%
50% of probability mass above
4
PLOS ONE
5266 papers in training set
Top 32%
4.4%
5
Frontiers in Medicine
120 papers in training set
Top 0.5%
4.4%
6
Artificial Intelligence in Medicine
17 papers in training set
Top 0.1%
3.3%
7
Patterns
78 papers in training set
Top 0.5%
3.3%
8
JAMIA Open
42 papers in training set
Top 0.5%
3.3%
9
Frontiers in Public Health
148 papers in training set
Top 2%
2.8%
10
Communications Medicine
113 papers in training set
Top 1%
2.7%
11
Cytometry Part A
33 papers in training set
Top 0.1%
2.2%
12
PLOS Computational Biology
1863 papers in training set
Top 14%
1.8%
13
Nature Communications
5641 papers in training set
Top 52%
1.1%
14
BMC Medical Informatics and Decision Making
43 papers in training set
Top 1%
1.1%
15
International Journal of Medical Informatics
26 papers in training set
Top 1%
1.1%
16
IEEE Access
35 papers in training set
Top 1%
1.1%
17
Nature Machine Intelligence
70 papers in training set
Top 2%
0.9%
18
Frontiers in Digital Health
24 papers in training set
Top 1%
0.9%
19
Advanced Science
286 papers in training set
Top 9%
0.9%
20
IEEE Journal of Biomedical and Health Informatics
37 papers in training set
Top 1%
0.6%
21
GigaScience
212 papers in training set
Top 5%
0.6%
22
npj Systems Biology and Applications
125 papers in training set
Top 2%
0.6%
23
JAMA Network Open
130 papers in training set
Top 4%
0.6%
24
Nature Medicine
125 papers in training set
Top 4%
0.6%