A Hybrid AutoML Ensemble Integrating Conventional Learners and Gradient-Boosting Models for Multi-Outcome Prediction in ICU Patients with Pseudomonas aeruginosa
Lv, X.-c.; Ren, Q.; Zhu, L.-h.; Chen, K.; Wang, J.-b.; Chen, F.; Jin, K.-l.; lin, k.
Show abstract
BackgroundCarbapenem resistance in Pseudomonas aeruginosa is increasing in intensive care units (ICUs). To enhance antimicrobial stewardship and infection control, we aimed to develop and validate a real-time interpretable hybrid Automated Machine Learning (AutoML) ensemble for multi-outcome prediction. MethodsWe retrospectively analyzed 847 adult ICU admissions with P. aeruginosa isol ates at a tertiary hospital in Hangzhou, China (January 2018 to December 2024). After a three-stage VTF-MI-L1 feature selection pipeline, XGBoost, LightGBM, CatBoost, random forests, and linear/logistic regression were used as base learners and combined via Bagging, Voting, Stacking, and Gradient Boosting. Nested five-fold cross-validation was used to assess model performance (AUC for classification; MSE, RMSE, MAE, and R2 for regression). Interpretability was provided by SHAP values, and the inference latency was recorded. ResultsFor carbapenem resistance rate (CRR) prediction, the CatBoost regressor (cRMSD = 0.1663; r = 0.8849; R2 {approx} 0.78) and the Voting Regressor (cRMSD = 0.1675; r = 0.8838) outperformed all other models (p < 0.05). XGB-R achieved the best accuracy and computational efficiency for the last two tests of the CRR of P. aeruginosa (CRR-PA-Last2) (p < 0.05). In predicting ICU length of stay, XGB-R led with r = 0.9724, cRMSD = 55.7 d, and {sigma}-ratio = 0.88, significantly surpassing Bagging and CatBoost regressors (p < 0.05). XGB-R also yielded the lowest composite error for the ICU-to-death interval (cRMSD = 205.1 d; r = 0.7741; {sigma}-ratio = 0.71), again outperforming Bagging and CatBoost (p < 0.05). Across all four regression outcomes, XGB-R obtained the lowest average rank (1.95), whereas CatBoost and Voting regressors showed particular strengths in predicting resistance. SHAP analysis identified age, carbapenem exposure intensity, and duration of mechanical ventilation and catheterization as the key positive contributors. All top-ranked models required < 50 ms per inference, meeting the bedside real-time requirements. ConclusionsThe proposed hybrid AutoML ensemble delivered highly accurate, interpretable, and millisecond-level predictions of diverse resistance-related outcomes, underscoring its potential for ICU antimicrobial stewardship and infection control. Multicenter prospective studies are warranted to confirm the generalizability of these findings.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Development of a Risk Prediction Model for Sepsis-Related Delirium Based on Multiple Machine Learning Approaches and an Online Calculator 95%
- A comparison of machine learning models versus clinical evaluation for mortality prediction in patients with sepsis 95%
- A Machine Learning-Based Prediction of Hospital Mortality in Mechanically Ventilated ICU Patients 95%
Similar papers in this journal
- Predicting the causative pathogen among children with pneumonia using a causal Bayesian network 93%
- Contrasting factors associated with COVID-19-related ICU admission and death outcomes in hospitalised patients by means of Shapley values 92%
- Bayesian modeling of the impact of antibiotic resistance on the efficiency of MRSA decolonization 92%
Similar papers in this journal
- Predicting bloodstream infection outcome using machine learning 96%
- Mitigating Machine Learning Bias Between High Income and Low-Middle Income Countries for Enhanced Model Fairness and Generalizability 94%
- Limitations of estimating antibiotic resistance using German hospital consumption data - A comprehensive computational analysis 93%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.