Early-Horizon Multimodal ICU Mortality Prediction Without Retraining
Bakumenko, A.; Smith, D. H.; Hoelscher, J.
Show abstract
Earlier ICU mortality prediction is more clinically useful because it can identify high-risk patients while treatment decisions can still change. Yet most models are trained on data from a fixed time window, so it is unclear whether a model trained on the first 48 hours of ICU data remains reliable when used earlier in the ICU stay. We evaluated a multimodal ICU mortality model trained once at 48 hours and then applied unchanged at 6, 12, 24, and 48 hours on MIMIC-III. The model combines an LSTM for physiological time-series data, a finetuned ClinicalModernBERT model for clinical notes, and a logistic regression fusion layer. Performance remained strong at earlier time points, suggesting that useful mortality prediction is possible earlier in the ICU stay even without retraining. At 6 hours, the model achieved AUROC 0.777 and remained well-calibrated (ECE 0.038) without any recalibration, and it outperformed both single-modality models at every horizon. The multimodal benefit was most evident at earlier horizons, when physiological data were sparse: agreement between the two specialists dropped by more than half from 48 to 6 hours, while the median contribution from clinical notes increased from 37% to 49%. A Bayesian version of the fusion layer showed that uncertainty decreased for survivors as more data accumulated but remained high for non-survivors; the most uncertain cases were up to 4.9 times more likely to be non-surviving patients. Continuous hourly analyses further showed that clinical notes provide stable context between documentation events. Simply carrying forward the most recent note matched or outperformed note-decay and documentation-gap alternatives. These results suggest that a multimodal ICU mortality model trained on 48 hours of data can provide trustworthy earlier predictions without retraining, while also identifying the cases that remain hardest to interpret.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- Emergency department admissions during COVID-19: explainable machine learning to characterise data drift and detect emergent health risks 95%
- Developing Machine Learning Models for Predicting Intensive Care Unit Resource Use During the COVID-19 Pandemic 95%
- EHR Foundation Models Improve Robustness in the Presence of Temporal Distribution Shift 94%
Similar papers in this journal
- Modular Clinical Decision Support Networks (MoDN)—Updatable, Interpretable, and Portable Predictions for Evolving Clinical Environments 95%
- Explainable deep learning for disease activity prediction in chronic inflammatory joint diseases 92%
- A flexible framework for minimal biomarker signature discovery from clinical omics studies without library size normalisation 91%
Similar papers in this journal
- Deep representation learning for clustering longitudinal survival data from electronic health records 92%
- Short-term forecasting of COVID-19 in Germany and Poland during the second wave – a preregistered study 92%
- Integrating T-cell receptor and transcriptome for large-scale single-cell immune profiling analysis 91%
Similar papers in this journal
- COSIME: Cooperative multi-view integration with Scalable and Interpretable Model Explainer 92%
- Estimating Treatment Effects for Time-to-Treatment Antibiotic Stewardship in Sepsis 92%
- Generalized Radiograph Representation Learning via Cross-supervision between Images and Free-text Radiology Reports 90%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.