Machine Learning for Urinary Tract Infection Prediction in Emergency Departments: An Explainable Approach
Donmez, A.
Show abstract
Urinary tract infections (UTIs) represent a substantial burden in emergency department (ED) settings, where diagnostic delays and the limitations of traditional clinical assessments often result in suboptimal treatment decisions. This study develops an interpretable machine learning framework to enhance real-time UTI prediction accuracy. We analyzed a retrospective dataset of 80,387 ED patient encounters from four institutions (2013-2016), encompassing 220 clinical variables. Four machine learning algorithms, Decision Tree, Random Forest, Logistic Regression, and XGBoost, were trained and evaluated. Model interpretability was achieved through SHapley Additive exPlanations (SHAP) analysis. Performance was assessed using area under the receiver operating characteristic curve (AUC), sensitivity, and specificity, both overall and across multiple patient subgroups. XGBoost demonstrated the highest overall performance with an AUC of 0.90 for general UTI prediction and 0.97 for complicated UTI identification. The reduced 20-feature XGBoost model achieved similar performance to the full feature model. SHAP analysis indicated that urinalysis markers, including leukocyte esterase, nitrites, and bacterial count, were the primary predictive features, with additional contributions from age, vital signs, and selected comorbidities. The model maintained robust performance across demographic and clinical subgroups, with AUC values ranging from 0.88 to 0.92. This study presents a clinically viable, explainable machine learning framework that addresses critical gaps in ED-based UTI diagnosis by combining high predictive accuracy with transparent feature attribution, reduced-feature implementation, and detailed subgroup and benchmark analyses.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Can the application of machine learning to electronic health records guide antibiotic prescribing decisions for suspected urinary tract infection in the Emergency Department? 96%
- Identification of predictive patient characteristics for assessing the probability of COVID-19 in-hospital mortality 94%
- Raising awareness of potential biases in medical machine learning: Experience from a Datathon 93%
Similar papers in this journal
Similar papers in this journal
- Development and Validation of ‘Patient Optimizer’ (POP) Algorithms for Predicting Surgical Risk with Machine Learning 95%
- Prediction of Sepsis Mortality in ICU Patients Using Machine Learning Methods 94%
- OASIS+: leveraging machine learning to improve the prognostic accuracy of OASIS severity score for predicting in-hospital mortality 94%
Similar papers in this journal
- Validation of a Derived International Patient Severity Algorithm to Support COVID-19 Analytics from Electronic Health Record Data 93%
- Development and Validation of Phenotype Classifiers across Multiple Sites in the Observational Health Sciences and Informatics (OHDSI) Network 93%
- Use of unstructured text in prognostic clinical prediction models: a systematic review 93%
Similar papers in this journal
- A comparison of machine learning models versus clinical evaluation for mortality prediction in patients with sepsis 95%
- A Machine Learning-Based Prediction of Hospital Mortality in Mechanically Ventilated ICU Patients 94%
- Real-world evidence: telemedicine for complicated cases of urinary tract infection 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.