An Interpretable Longitudinal Preeclampsia Risk Prediction Using Machine Learning
Eberhard, B. W.; Cohen, R. Y.; Rigoni, J.; Bates, D.; Gray, K. J.; Kovacheva, V.
Show abstract
BackgroundPreeclampsia is a pregnancy-specific disease characterized by new onset hypertension after 20 weeks of gestation that affects 2-8% of all pregnancies and contributes to up to 26% of maternal deaths. Despite extensive clinical research, current predictive tools fail to identify up to 66% of patients who will develop preeclampsia. We sought to develop a tool to longitudinally predict preeclampsia risk. MethodsIn this retrospective model development and validation study, we examined a large cohort of patients who delivered at six community and two tertiary care hospitals in the New England region between 02/2015 and 06/2023. We used sociodemographic, clinical diagnoses, family history, laboratory, and vital signs data. We developed eight datasets at 14, 20, 24, 28, 32, 36, 39 weeks gestation and at the hospital admission for delivery. We created linear regression, random forest, xgboost, and deep neural networks to develop multiple models and compared their performance. We used Shapley values to investigate the global and local explainability of the models and the relationships between the predictive variables. FindingsOur study population (N=120,752) had an incidence of preeclampsia of 5.7% (N=6,920). The performance of the models as measured using the area under the curve, AUC, was in the range 0.73-0.91, which was externally validated. The relationships between some of the variables were complex and non-linear; in addition, the relative significance of the predictors varied over the pregnancy. Compared to the current standard of care for preeclampsia risk stratification in the first trimester, our model would allow 48.6% more at-risk patients to be identified. InterpretationOur novel preeclampsia prediction tool would allow clinicians to identify patients at risk early and provide personalized predictions, as well as longitudinal predictions throughout pregnancy. FundingNational Institutes of Health, Anesthesia Patient Safety Foundation. RESEARCH IN CONTEXTO_ST_ABSEvidence before this studyC_ST_ABSCurrent tools for the prediction of preeclampsia are lacking as they fail to identify up to 66% of the patients who develop preeclampsia. We searched PubMed, MEDLINE, and the Web of Science from database inception to May 1, 2023, using the keywords "deep learning", "machine learning", "preeclampsia", "artificial intelligence", "pregnancy complications", and "predictive models". We identified 13 studies that employed machine learning to develop prediction models for preeclampsia risk based on clinical variables. Among these studies, six included biomarkers such as serum placental growth factor, pregnancy-associated plasma protein A, and uterine artery pulsatility index, which are not routinely available in our clinical practice; two studies were in diverse cohorts of more than 100 000 patients, and two studies developed longitudinal predictions using medical records data. However, most studies have limited depth, concerns about data leakage, overfitting, or lack of generalizability. Added value of this studyWe developed a comprehensive longitudinal predictive tool based on routine clinical data that can be used throughout pregnancy to predict the risk of preeclampsia. We tested multiple types of predictive models, including machine learning and deep learning models, and demonstrated high predictive power. We investigated the changes over different time points of individual and group variables and found previously known and novel relationships between variables such as red blood cell count and preeclampsia risk. Implications of all the available evidenceLongitudinal prediction of preeclampsia using machine learning can be achieved with high performance. Implementation of an accurate predictive tool within the electronic health records can aid clinical care and identify patients at heightened risk who would benefit from aspirin prophylaxis, increased surveillance, early diagnosis, and escalation in care. These results highlight the potential of using artificial intelligence in clinical decision support, with the ultimate goal of reducing iatrogenic preterm birth and improving perinatal care.
Matching journals
The top 9 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Predictability and Stability Testing to Assess Clinical Decision Instrument Performance for Children After Blunt Torso Trauma 92%
- A proposed de-identification framework for a cohort of children presenting at a health facility in Uganda 91%
- Generalizability Challenges of Mortality Risk Prediction Models: A Retrospective Analysis on a Multi-center Database 90%
Similar papers in this journal
- Who is pregnant? defining real-world data-based pregnancy episodes in the National COVID Cohort Collaborative (N3C) 93%
- Development and Application of Pharmacological Statin-Associated Muscle Symptoms Phenotyping Algorithms Using Structured and Unstructured Electronic Health Records Data 91%
- Clinical Study Applying Machine Learning to Detect a Rare Disease: Results and Lessons Learned 90%
Similar papers in this journal
- Widely accessible prognostication using medical history for fetal growth restriction and small for gestational age in nationwide insured women 96%
- Preeclampsia prediction with maternal and paternal polygenic risk scores: the TMM BirThree Cohort Study 95%
- Genetic Predictors of Blood Pressure Traits are Associated with Preeclampsia 95%
Similar papers in this journal
- A Deep Learning Approach for Transgender and Gender Diverse Patient Identification in Electronic Health Records 90%
- Natural language processing for scalable feature engineering and ultra-high-dimensional confounding adjustment in healthcare database studies 89%
- Automated Interpretable Discovery of Heterogeneous Treatment Effectiveness: A Covid-19 Case Study 89%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.