Machine learning prediction of early postpartum prediabetes in women with gestational diabetes mellitus
Parkhi, D. M.; Periyathambi, N.; Weldeselassie, Y.; Patel, V.; Sukumar, N.; Siddharthan, R.; Narlikar, L.; Saravanan, P.
Show abstract
BackgroundEarly onset of type 2 diabetes and cardiovascular disease are common complications for women diagnosed with gestational diabetes. About half of the women with gestational diabetes develop postpartum prediabetes within 10 years of the index pregnancy. These women also have double the risk of developing cardiovascular disease than women without a history of gestational diabetes. Currently, there is no accurate way of knowing which women with gestational diabetes are likely to develop postpartum prediabetes. This study aims to predict the risk of postpartum prediabetes in women diagnosed with gestational diabetes. MethodsWe build a sparse logistic regression-based machine learning model to learn key variables significant for the prediction of postpartum prediabetes, from antenatal data with maternal anthropometric and biochemical variables as well as neonatal characteristics of 607 UK women diagnosed with gestational diabetes. We evaluate the performance of the proposed model in addition to other more advanced machine learning methods using established metrics such as the area under the receiver operating characteristic curve and specificity for pre-determined values of sensitivity. We use K-L divergence and information graphs to evaluate and compare different thresholds of classification for targeted screening options in resource-constrained settings. We also perform a decision curve analysis to study the net standardized benefit of our model compared to the universal screening approach. ResultsStrikingly, our sparse logistic regression approach selects only two variables as relevant but gives an area under the receiver operating characteristic curve of 0.72, outperforming all other methods. It can identify postpartum prediabetes in women with gestational diabetes using the Rule-in test with 92% specificity at an optimal probability threshold of 0.381 and using the Rule-out test with 92% sensitivity at an optimal probability threshold of 0.140. ConclusionWe propose a simple logistic regression model, which needs only the antenatal fasting glucose at OGTT and HbA1c soon after the diagnosis of GDM, to predict, with remarkable accuracy, the probability of postpartum prediabetes in women with gestational diabetes. We envision this to be a practical solution, which coupled with a targeted follow-up of high-risk women, could yield better cardiometabolic outcomes in women with a history of GDM.
Matching journals
The top 8 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Estimating Heritability of Glycaemic Response to Metformin using Nationwide Electronic Health Records and Population-Sized Pedigree 91%
- Precision Gestational Diabetes Treatment: Systematic review and Meta-analyses 90%
- Thyroid dysfunction diagnosis from routine laboratory tests based on machine learning 90%
Similar papers in this journal
- Widely accessible prognostication using medical history for fetal growth restriction and small for gestational age in nationwide insured women 95%
- Can machine learning improve risk prediction of incident hypertension? An internal method comparison and external validation of the Framingham risk model using HUNT Study data 94%
- Pregnancy-induced changes in blood composition drive post-partum hemorrhage risk 92%
Similar papers in this journal
- Predicting long-term Type 2 Diabetes with Support Vector Machine using Oral Glucose Tolerance Test 94%
- Early postpartum HbA1c after hyperglycemia first detected in pregnancy - imperfect but not without value 94%
- Optimization of nutritional strategies using a mechanistic computational model in prediabetes: Application to the J-DOIT1 study data 94%
Similar papers in this journal
- Individual Reference Intervals for Personalized Interpretation of Clinical and Metabolomics Measurements 93%
- A methodology of phenotyping ICU patients from EHR data: high-fidelity, personalized, and interpretable phenotypes estimation 92%
- Computational Strategies in Nutrigenetics: Constructing a Reference Dataset of Nutrition-Associated Genetic Polymorphisms 92%
Similar papers in this journal
- Racial disparities in continuous glucose monitoring-based 60-min glucose predictions among people with type 1 diabetes 95%
- Identification of predictive patient characteristics for assessing the probability of COVID-19 in-hospital mortality 91%
- Automated Image Transcription for Perinatal Blood Pressure Monitoring Using Mobile Health Technology 91%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.