Back

Predicting Inpatient Risk of Mortality in Diabetic Patients Using Administrative Data and Machine Learning: An External Validation Study Using SPARCS

Mirza, A. F.; Nwokeji, T. U.

2025-06-06 health informatics
10.1101/2025.06.05.25329076 medRxiv
Show abstract

ObjectivesTo evaluate whether machine learning models trained solely on administrative and demographic data can predict inpatient APR Risk of Mortality in diabetic patients. DesignRetrospective cohort study using New York State SPARCS data from 2021 and 2022. SettingNew York Statewide Planning and Research Cooperative System (SPARCS) data from 2021 and 2022. ParticipantsAdult inpatient admissions (age [≥]18) with a diagnosis of diabetes mellitus. Primary outcome measureAPR-DRG Risk of Mortality (ROM), classified as Minor, Moderate, Major, or Extreme. ResultsXGBoost outperformed logistic regression and random forest across all metrics. On the 2022 validation set, XGBoost achieved the highest accuracy (46.5%), macro AUC (0.699), weighted F1-score (0.458), and the lowest Brier score for the Extreme class (0.052). SHAP analysis identified length of stay, age group, and payer type as key predictors. ConclusionsEven without clinical data, administrative features contain non-random signals relevant for mortality risk stratification. These models, especially XGBoost, may help hospitals flag high-risk patients early using routinely available data, aiding triage and planning before labs or vitals are available. Strengths and limitations of this studyO_LIThis study is one of the first to apply machine learning to publicly available SPARCS data to predict APR-DRG Risk of Mortality in diabetic inpatients. C_LIO_LIWe evaluated three models using temporally distinct training and validation cohorts, simulating real-world model deployment across calendar years. C_LIO_LIModel interpretability was addressed using SHAP, providing transparent insights into feature contributions and enabling clinician-facing explanation. C_LIO_LIThe models relied solely on administrative and demographic data, limiting predictive fidelity due to the absence of clinical features such as laboratory values or vital signs. C_LIO_LIRisk of Mortality labels were derived from APR-DRG software and may be influenced by coding practices rather than objective clinical outcomes. C_LI

Matching journals

The top 4 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.