Bayesian Prediction of Severe Outcomes in the LabMarCS: Laboratory Markers of COVID-19 Severity - Bristol Cohort
Sullivan, B.; Barker, E.; Williams, P.; MacGregor, L.; Bhamber, R.; Thomas, M.; Gurney, S.; Hyams, C.; Whiteway, A.; Cooper, J. A.; McWilliams, C.; Turner, K.; Dowsey, A. W.; Albur, M.
Show abstract
We describe several regression models to predict severe outcomes in COVID-19 and challenges present in complex observational medical data. We demonstrate best practices for data curation, cross-validated statistical modelling, and variable selection emphasizing recent Bayesian methods. The study follows a retrospective observational cohort design using multicentre records across National Health Service (NHS) trusts in southwest England, UK. Participants included hospitalised adult patients positive for SARS-CoV 2 during March to October 2020, totalling 843 patients (mean age 71, 45% female, 32% died or needed ICU stay), split into training (n=590) and validation groups (n=253). Models were fit to predict severe outcomes (ICU admission or death within 28-days of admission to hospital for COVID-19, or a positive PCR result if already admitted) using demographic data and initial results from 30 biomarker tests collected within 3 days of admission or testing positive if already admitted. Cross-validation results showed standard logistic regression had an internal validation median AUC of 0.74 (95% Interval [0.62,0.83]), and external validation AUC of 0.68 [0.61, 0.71]; a Bayesian logistic regression (with horseshoe prior) internal AUC of 0.79 [0.71, 0.87], and external AUC of 0.70 [0.68, 0.71]. Variable selection performed using Bayesian predictive projection determined a four variable model using Age, Urea, Prothrombin time and Neutrophil-Lymphocyte ratio, with a median internal AUC of 0.79 [0.78, 0.80], and external AUC of 0.67 [0.65, 0.69]. We illustrate best-practices protocol for conventional and Bayesian prediction modelling on complex clinical data and reiterate the predictive value of previously identified biomarkers for COVID-19 severity assessment.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- A comparison of machine learning models versus clinical evaluation for mortality prediction in patients with sepsis 95%
- An Online Risk Calculator for Rapid Prediction of In-hospital Mortality from COVID-19 Infection 94%
- A Machine Learning-Based Prediction of Hospital Mortality in Mechanically Ventilated ICU Patients 94%
Similar papers in this journal
- Modeling physician variability to prioritize relevant medical record information 95%
- A deep learning model for clinical outcome prediction using longitudinal inpatient electronic health records 95%
- Characterizing subgroup performance of probabilistic phenotype algorithms within older adults: A case study for dementia, mild cognitive impairment, and Alzheimer’s and Parkinson’s diseases 93%
Similar papers in this journal
- Towards reduction in bias in epidemic curves due to outcome misclassification through Bayesian analysis of time-series of laboratory test results: Case study of COVID-19 in Alberta, Canada and Philadelphia, USA 93%
- Urinary tract infections in children: building a causal model-based decision support tool for diagnosis with domain knowledge and prospective data 93%
- Predictive accuracy of a hierarchical logistic model of cumulative SARS-CoV-2 case growth 91%
Similar papers in this journal
- Predicting the need for escalation of care or death from repeated daily clinical observations and laboratory results in patients with SARS-CoV-2 during 2020: a retrospective population-based cohort study from the United Kingdom 94%
- Performance of Existing and Novel Symptom- and Antigen Testing-Based COVID-19 Case Definitions in a Community Setting 93%
- Obtaining prevalence estimates of COVID-19: A model to inform decision-making 92%
Similar papers in this journal
- Predicting bloodstream infection outcome using machine learning 96%
- Mitigating Machine Learning Bias Between High Income and Low-Middle Income Countries for Enhanced Model Fairness and Generalizability 95%
- Developing And Validating COVID-19 Adverse Outcome Risk Prediction Models From A Bi-National European Cohort Of 5594 Patients 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.