Fairness-Aware Machine Learning for Heart Failure Prediction: Performance, Bias, and Clinical Deployment Insights
Eziama, E. U.; Eziama, S. I.; Agboli, V. I.; Amusa, T.; Edokpolor, H.; Adeoba, M. I.; Usman, I. A.; Eghujovbo, V.; Olasege, I.; Ogunnaike, O.; Ige, T. O.; Okunola, D.; Taiwo, F. T.; Anyanacho, J. N.; Olasege, R.; Adeoye, A.
Show abstract
Heart failure (HF) prediction models using machine learning (ML) must achieve a balance between performance, fairness, and real-world clinical utility. This paper assesses the potential of ML and DL models in the context of heterogeneous databases (UCI, MIMIC) and aims to derive applicable schemes for equitable deployment in healthcare. Although the Transformer models depicted notable AUC-ROC in the UCI data (0.986), they suffered a considerable performance degradation in the MIMIC data, suggesting their generalizability issue. Interestingly, we found contrasting gender preferences: SVM better detected females (AUC-ROC = 0.869 vs 0.796 for males), while XGBoost favored males. Decision curve analysis (DCA) demonstrated that a greater AUC-ROC was not necessarily preferred. For the 0.5 threshold, the net benefit of SVM (0.050) was higher than that of other models. SHAP analysis confirmed sex, ejection fraction, and NT-proBNP as the main predictors, but demonstrated inconsistent feature relations between datasets. We argue that (1) gender-specific threshold optimization, (2)ensemble methods to counteract bias, and (3)robust external validation are necessary and provide a possible path for clinically deployable, fairness-aware HF prediction tools.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Identification of predictive patient characteristics for assessing the probability of COVID-19 in-hospital mortality 94%
- From theoretical models to practical deployment: A perspective and case study of opportunities and challenges in AI-driven healthcare research for low-income settings 94%
- QRS detection in single-lead, telehealth electrocardiogram signals: benchmarking open-source algorithms 94%
Similar papers in this journal
- A Machine Learning-Based Prediction of Hospital Mortality in Mechanically Ventilated ICU Patients 96%
- Predicting 30-Day and 1-Year Mortality in Heart Failure with Preserved Ejection Fraction (HFpEF) 95%
- Enhanced machine learning and hybrid ensemble approaches for coronary heart disease prediction 94%
Similar papers in this journal
- Detecting Heart Failure using novel bio-signals and a Knowledge Enhanced Neural Network 95%
- Improving irregular temporal modeling by integrating synthetic data to the electronic medical record using conditional GANs: a case study of fluid overload prediction in the intensive care unit 95%
- Identification of Myocardial Infarction (MI) Probability from Imbalanced Medical Survey Data: An Artificial Neural Network (ANN) with Explainable AI (XAI) Insights 93%
Similar papers in this journal
- Towards Clinical Prediction with Transparency: An Explainable AI Approach to Survival Modelling in Residential Aged Care 94%
- A standardized analytics pipeline for reliable and rapid development and validation of prediction models using observational health data 93%
- Improving Heart Disease Probability Prediction Sensitivity with a Grow Network Model 93%
Similar papers in this journal
- Prediction of Unplanned 30- day Readmission for ICU Patients with Heart Failure 96%
- OASIS+: leveraging machine learning to improve the prognostic accuracy of OASIS severity score for predicting in-hospital mortality 96%
- Optimized Feature Selection and Advanced Machine Learning for Stroke Risk Prediction in Revascularized Coronary Artery Disease Patients 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.