Predicting ICU Transfer and Short-term Mortality in Emergency Department Atrial Fibrillation Patients: An Enhanced Machine Learning Model Using MIMIC Data
Pishgar, M.; Li, T.; Li, Z.; Chen, S.; Pishgar, E.; Alaei, K.; Placencia, G.
Show abstract
Atrial fibrillation (AF) is a prevalent condition in emergency department (ED) patients and is associated with an elevated risk of intensive care unit (ICU) transfer and short-term mortality. Early identification of high-risk patients is critical for timely intervention and improved clinical outcomes. We constructed a combined cohort from the MIMIC-IV-ED and MIMIC-IV databases, comprising ED admissions of patients with AF, and developed an interpretable machine learning (ML) framework to predict ICU transfer and mortality within 3, 7, and 30 days using clinical variables obtained at triage. The preprocessing pipeline included imputation of missing data, z-score normalization of numerical features, one-hot encoding of categorical variables, and correction for class imbalance using the Synthetic Minority Over-sampling Technique (SMOTE). A hybrid feature selection strategy combining Recursive Feature Elimination with Cross-Validation (RFECV) and Least Absolute Shrinkage and Selection Operator (LASSO) regression reduced the initial set to 19 clinically relevant predictors. Among six evaluated machine learning algorithms, LightGBM demonstrated the highest performance for ICU transfer (AUROC = 0.7979, 95% confidence interval (CI): 0.7916-0.8041) and for 7- and 30-day mortality (AUROC = 0.8316, 95% CI: 0.8156-0.8476; AUROC = 0.8010, 95% CI: 0.7898-0.8123), while CatBoost achieved the best performance for 3-day mortality (AUROC = 0.8444, 95% CI: 0.8237-0.8644). SHAP(SHapley Additive exPlanations) analysis identified O2sat, acuity, and resprate as key determinants, underscoring the clinical plausibility and interpretability of the models. These findings highlight the potential of interpretable machine learning approaches to enable early, time-sensitive risk stratification and support informed clinical decision-making for AF patients in the ED.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- QRS detection in single-lead, telehealth electrocardiogram signals: benchmarking open-source algorithms 94%
- Detecting QT prolongation From a Single-lead ECG With Deep Learning 94%
- Identification of predictive patient characteristics for assessing the probability of COVID-19 in-hospital mortality 92%
Similar papers in this journal
- Identification of Digital Twins to Guide Interpretable AI for Diagnosis and Prognosis in Heart Failure 93%
- AI Learning for Pediatric Right Ventricular Assessment: Development and Validation Across Multiple Centers 93%
- Cohort Design and Natural Language Processing to Reduce Bias in Electronic Health Records Research: The Community Care Cohort Project 92%
Similar papers in this journal
- The Transcriptional Landscape of Atrial Fibrillation: A Systematic Review and Meta-analysis 94%
- An assessment of the value of deep neural networks in genetic risk prediction for surgically relevant outcomes 94%
- Predicting 30-Day and 1-Year Mortality in Heart Failure with Preserved Ejection Fraction (HFpEF) 94%
Similar papers in this journal
- OASIS+: leveraging machine learning to improve the prognostic accuracy of OASIS severity score for predicting in-hospital mortality 94%
- Optimized Feature Selection and Advanced Machine Learning for Stroke Risk Prediction in Revascularized Coronary Artery Disease Patients 94%
- Prediction of Unplanned 30- day Readmission for ICU Patients with Heart Failure 93%
Similar papers in this journal
- Predicting bloodstream infection outcome using machine learning 93%
- Advancing Cardiovascular Disease Diagnosis: A Robust ML Ecosystem Integrating Early Detection, Responsible AI Framework, and Causal Inference 93%
- Explaining Deep Neural Networks for Knowledge Discovery in Electrocardiogram Analysis 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.