Predicting Hypertension Among HIV Patients on Antiretroviral Therapy in Rural Eastern Cape, South Africa Using Machine Learning
Tsuro, U.; Ncube, T.; Oladimeji, K. E.; Apalata, T. R.
Show abstract
BackgroundHypertension continues to be a major challenge in developing countries like South Africa, as it significantly contributes to the cardiovascular disease burden in these countries. This study aimed to utilize the machine learning (ML) models to anticipate the incidence of hypertension in HIV patients under antiretroviral therapy (ART) in rural Eastern Cape, South Africa. MethodsThis research carried out a retrospective cohort study and created and tested six machine learning algorithms: Neural Networks, Random Forest, Logistic Regression, Naive Bayes, K-Nearest Neighbours and XGBoost. The goal was to predict the likelihood of developing hypertension. Feature selection was done using the Boruta method and the model was assessed using several metrics including aiming, precision, recall, F1 score, and area under the receiver operating characteristic curve (AUC). ResultsXGBoost outperformed all other models with an AUC of 0.96, which further suggests it can effectively distinguish between hypertensives and normotensives. In the case of Boruta analysis, some aggravated risk factors were age category, time on ART, BMI category, waist to hip ratio, waist size, family history of HBP and relationship status, physical activity, LDL cholesterol level, awareness of high blood pressure, education level, use of ART and diabetes mellitus. ConclusionsThis study has highlighted the utility of XGBoost, as one of the advanced machine learning algorithms, in reliably forecasting the occurrence of hypertension in HIV ART patients in a rural setting. The established risk factors elucidate the complexity behind the hypertension emergence and hence the need for triad approaches which include lifestyle changes, clinical treatments, and demographic solutions to tackle the public health problem.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Undiagnosed hypertension and its associated factors in India: A rural-urban contrast from the National Family Health Survey (2019-21) 96%
- Prevalence and determinants of peripheral arterial disease in children with nephrotic syndrome 96%
- Comparison of WHO laboratory-based and non-laboratory-based CVD risk Charts among Hypertensive Adults Attending Primary Healthcare Centers in West Africa Sub-region 96%
Similar papers in this journal
- An AI-based approach to predict delivery outcome based on measurable factors of pregnant mothers 96%
- Identification of predictive patient characteristics for assessing the probability of COVID-19 in-hospital mortality 94%
- Automated Image Transcription for Perinatal Blood Pressure Monitoring Using Mobile Health Technology 94%
Similar papers in this journal
- Unsupervised Discovery of Risk Profiles on Negative and Positive COVID-19 Hospitalized Patients 97%
- Identification of Myocardial Infarction (MI) Probability from Imbalanced Medical Survey Data: An Artificial Neural Network (ANN) with Explainable AI (XAI) Insights 96%
- A machine-learning Approach for Stress Detection Using Wearable Sensors in Free-living Environments 94%
Similar papers in this journal
- Classification models for Invasive Ductal Carcinoma Progression, based on gene expression data-trained supervised machine learning 95%
- A multipurpose machine learning approach to predict COVID-19 negative prognosis in Sao Paulo, Brazil 94%
- Rapid Clinical Screening and Staging for COVID-19 Severe Outcome A Hospitalization Study in New York City 94%
Similar papers in this journal
- Prediction of Sepsis Mortality in ICU Patients Using Machine Learning Methods 95%
- Combining symbolic regression with the Cox proportional hazards model improves prediction of heart failure deaths 94%
- On the predictability of postoperative complications for cancer patients: a Portuguese cohort study 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.