Identifying and ranking novel independent features for cardiovascular disease prediction in people with type 2 diabetes
Dziopa, K.; Chaturvedi, N.; Asselbergs, F.; Schmidt, A. F.
Show abstract
BackgroundCVD prediction models do not perform well in people with diabetes. We therefore aimed to identify novel predictors for six facets of CVD, (including coronary heart disease (CHD), Ischemic stroke, heart failure (HF), and atrial fibrillation (AF)) in people with T2DM. MethodsAnalyses were conducted using the UK biobank and were stratified on history of CVD and of T2DM: 459,142 participants without diabetes or a history of CVD, 14,610 with diabetes but without CVD, and 4,432 with diabetes and a history of CVD. Replication was performed using a 20% hold-out set, ranking features on their permuted c-statistic. ResultsOut of the 600+ candidate features, we identified a subset of replicated features, ranging between 32 for CHD in people with diabetes to 184 for CVD+HF+AF in people without diabetes. Classical CVD risk factors (e.g. parental or maternal history of heart disease, or blood pressure) were relatively highly ranked for people without diabetes. The top predictors in the people with diabetes without a CVD history included: cystatin C, self-reported health satisfaction, biochemical measures of ill health (e.g. plasma albumin). For people with diabetes and a history of CVD top features were: self-reported ill health, and blood cell counts measurements (e.g. red cell distribution width). We additionally identified risk factors unique to people with diabetes, consisting of information on dietary patterns, mental health and biochemistry measures. Consideration of these novel features improved risk classification, for example per 1000 people with diabetes 133 CVD and 165 HF cases appropriately received a higher risk. ConclusionThrough data-driven feature selection we identified a substantial number of features relevant for prediction of cardiovascular risk in people with diabetes, the majority of which related to non-classical risk factors such as mental health, general illness markers, and kidney disease.
Matching journals
The top 9 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Cardiovascular autonomic dysfunction precedes cardiovascular disease and all-cause mortality: 11-year follow-up of the ADDITION-PRO study 94%
- Heterogeneity of Treatment Effects Across Nine Glucose-Lowering Drug Classes in Type 2 Diabetes: Extension of the LEGEND-T2DM Network Study 94%
- Covid-19 fatality prediction in people with diabetes and prediabetes using a simple score at hospital admission 93%
Similar papers in this journal
- Development and validation of a risk prediction algorithm for high-risk populations combining genetic and conventional risk factors of cardiovascular disease 93%
- Longitudinal associations of sustained low or high income and income variability with incident cardiovascular disease in individuals with type 2 diabetes: a retrospective population-based cohort study 93%
- Artificial Intelligence Methods to Detect Heart Failure with Preserved Ejection Fraction (AIM-HFpEF) within Electronic Health Records: An equitable disease prediction model 91%
Similar papers in this journal
Similar papers in this journal
- Cardiovascular risk prediction in type 2 diabetes: a comparison of 22 risk scores in primary care setting 98%
- Discovery of biomarkers for glycaemic deterioration before and after the onset of type 2 diabetes: an overview of the data from the epidemiological studies within the IMI DIRECT Consortium 92%
- Depression, diabetes, their comorbidity and all-cause and cause-specific mortality: a prospective cohort study 92%
Similar papers in this journal
- Precision Prognostics for Cardiovascular Disease in Type 2 Diabetes: A Systematic Review and Meta-analysis 95%
- Systematic review of precision subclassification of type 2 diabetes 92%
- Estimating Heritability of Glycaemic Response to Metformin using Nationwide Electronic Health Records and Population-Sized Pedigree 90%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.