Predicting Hospital Admissions Using Pretrained EHR Embeddings: External Evaluation and Insights on Local Vocabulary Adaptation
Neves, B.; Cerejo, J.; Goncalves, S.; Mota, I.; Moreira, J. M.; Silva, N. A.; Leite, F.; Wornow, M.; Silva, M. J.
Show abstract
PurposeUnplanned hospital admissions impose substantial strain on healthcare systems, yet predictive models for these events remain underexplored in practice. This study evaluates whether publicly available pretrained transformer-based embeddings, developed on an external health system, can improve prediction of hospital admissions--including unplanned cases--when applied to a different institution with sparser data and a distinct medical vocabulary. MethodsWe performed a retrospective cohort study using structured EHR data from 200,000 adult patients (2007-2023) at a Portuguese hospital, standardized to the OMOP Common Data Model. Four 30-day outcomes were predicted: emergency department visits, hospital admissions, unplanned admissions, and readmissions. Three modeling approaches were compared: (1) clinically curated handcrafted features, (2) frequency-based representations of all recorded OMOP concepts, and (3) pretrained CLMBR-T embeddings generated from longitudinal OMOP data of 2.57 million patients in a U.S. hospital system. Performance was assessed on held-out patients using AUROC, AUPRC, and calibration metrics, with additional analysis of the impact of vocabulary overlap between pretraining and local datasets. ResultsPretrained embeddings achieved the highest discrimination for all outcomes, particularly for unplanned admissions (AUROC 0.877 vs. 0.770 for counts). Gains were greatest for rarer outcomes and patients with richer clinical histories. Despite only 58% overlap with local vocabulary and substantially fewer events per patient than in pretraining, embeddings transferred effectively, indicating generalizable temporal patterns. Calibration was poorer than simpler models, necessitating post-hoc recalibration before deployment. ConclusionPretrained OMOP-based EHR embeddings can substantially improve prediction of hospital and unplanned admissions in data- and resource-limited settings, even with partial vocabulary overlap. These findings support their use for rapid, cost-effective deployment of clinically meaningful predictive models, provided local recalibration and workflow integration are addressed. HighlightsO_LISuperior cross-institution performance - Pretrained EHR embeddings from Stanford Medicine achieved AUROC 0.877 for unplanned admissions, 0.814 for hospital admissions, 0.782 for ED visits, and 0.923 for readmissions in a Portuguese hospital, outperforming count-based (0.770, 0.767, 0.744, 0.922) and handcrafted feature models across all tasks. C_LIO_LILargest gains for rare, unpredictable events - For unplanned admissions (0.4% prevalence), embeddings nearly tripled AUPRC compared to counts (0.037 vs. 0.011) and improved AUROC by 0.107, with performance continuing to scale with more training data, unlike baselines. C_LIO_LIEffective under substantial domain shift - Strong transferability observed despite only 58% vocabulary overlap and markedly different patient populations, coding distributions, and event density (707 vs. 71 events per patient). C_LIO_LIBenefit increases with richer patient histories - Performance advantage of embeddings widened in patients with higher code volumes; AUROC for hospital admissions rose from 0.694 in the lowest quartile (Q1) to 0.870 in the highest (Q4), outperforming counts by up to 0.105 in Q4. C_LIO_LIActionable guidance for adoption - Hospitals with sparse data can achieve rapid, cost-effective deployment of predictive models using external embeddings, especially for rare outcomes, if paired with local fine-tuning and post-hoc recalibration to ensure accurate risk estimation before clinical use. C_LI
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
- EHR Foundation Models Improve Robustness in the Presence of Temporal Distribution Shift 97%
- Emergency department admissions during COVID-19: explainable machine learning to characterise data drift and detect emergent health risks 95%
- Leveraging Temporal Learning with Dynamic Range (TLDR) for Enhanced Prediction of Outcomes in Recurrent Exposure and Treatment Settings in Electronic Health Records 95%
Similar papers in this journal
- A Deep Learning Method to Detect Opioid Prescription and Opioid Use Disorder from Electronic Health Records 94%
- Image and structured data analysis for prognostication of health outcomes in patients presenting to the Emergency Department during the COVID-19 pandemic 93%
- Assessing the effects of data drift on the performance of machine learning models used in clinical sepsis prediction 92%
Similar papers in this journal
- Addressing Label Noise for Electronic Health Records: Insights from Computer Vision for Tabular Data 95%
- Implicit bias in Critical Care Data: Factors affecting sampling frequencies and missingness patterns of clinical and biological variables in ICU Patients 93%
- An Interpretable Risk Prediction Model for Healthcare with Pattern Attention 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.