Repurposing cardiovascular disease prediction models for cancer
Quill, S.; Hingorani, A. D.; Chaturvedi, N.; Schmidt, A. F.
Show abstract
BackgroundPopulation cancer screening detects the presence of early-stage disease rather than assessing future disease risk. We evaluated whether widely implemented cardiovascular disease (CVD) risk models can predict 10-year cancer risk and compared them with a less widely used cancer risk model (QCancer). MethodsWe evaluated four CVD prediction models: QRISK3, the Pooled Cohort Equations (PCE), SCORE2 and SCORE2-OP. All models were recalibrated using 20% of the UK Biobank (UKB) cohort and tested in the remainder, as well as in the Clinical Practice Research Datalink (CPRD). We gauged model performance using c-statistics for discrimination and evaluated the fidelity of calibration. We also identified the most influential risk factors in the QRISK3 model. FindingsIn the UKB test set, the c-statistics for incident CVD ranged from 0{middle dot}71 to 0{middle dot}74 (11,022 events). All CVD models achieved a c-statistic of 0{middle dot}63 for any cancer (23,010 events) and showed CVD-equivalent discrimination for gastro-oesophageal, liver and biliary tree, laryngeal, renal tract, and lung cancers (c-statistic range: 0{middle dot}70;0{middle dot}81). Overall, the discrimination of the CVD models was comparable that of the QCancer models (median difference in c-statistic: -0{middle dot}01 (95%CI -0{middle dot}03;0{middle dot}00). The recalibrated CVD models showed near-perfect calibration (median intercept 0{middle dot}01, Q1;Q3 -0{middle dot}05;0{middle dot}03 and slope 1{middle dot}00, Q1;Q3 0{middle dot}93;1{middle dot}15). Performance in CPRD (393,658 cancer events) was similar: the median c-statistic, calibration intercept, and slope were 0{middle dot}01 (95%CI 0{middle dot}00;0{middle dot}02), 0{middle dot}05 (95%CI 0{middle dot}02;0{middle dot}17), and 0{middle dot}04 (95%CI 0{middle dot}01;0{middle dot}15) higher, respectively, in CPRD than in UKB. After age, smoking status and systolic blood pressure were the most influential predictors of cancer risk. InterpretationWidely implemented CVD prediction models perform similarly to the QCancer models in the prediction of incident cancers. They may be used to inform cancer prevention and guide risk-stratified monitoring. The recalibrated models are available through an API. FundingHealth Data Research UK, British Heart Foundation and UK Research and Innovation.
Matching journals
The top 9 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Cancer and the risk of COVID-19 diagnosis, hospitalisation, and death: a population-based multi-state cohort study including 4,618,377 adults in Catalonia, Spain 95%
- Effects of Screening for Colorectal Cancer: Development, Documentation and Validation of a Multistate Markov Model 93%
- Free testosterone and malignant melanoma risk in men: prospective analyses of testosterone and SHBG with 19 cancers in men and postmenopausal women UK Biobank 93%
Similar papers in this journal
- Factors associated with excess all-cause mortality in the first wave of COVID-19 pandemic in the UK: a time-series analysis using the Clinical Practice Research Datalink 93%
- Genetically-proxied therapeutic inhibition of antihypertensive drug targets and risk of common cancers 93%
- Sociodemographic Characteristics and Longitudinal Progression of Multimorbidity: A Multistate Modelling Analysis of a Large Primary Care Records Dataset in England 91%
Similar papers in this journal
- An external validation of the QCovid risk prediction algorithm for risk of mortality from COVID-19 in adults: national validation cohort study in England 93%
- Differences in estimates for ten-year risk of cardiovascular disease in Black versus white persons with identical risk factor profiles using pooled cohort equations 92%
- Predictive performance and clinical application of COV50, a urinary proteomic biomarker in early COVID-19 infection: a cohort study 91%
Similar papers in this journal
- Healthy lifestyle, genetic risk, and incidence of cancer: A prospective cohort study of 13 cancer types 95%
- Causal relationships between risk of venous thromboembolism and 18 cancers: a bidirectional Mendelian randomisation analysis 93%
- Polygenic Risk Scores for Prediction of Breast Cancer in Korean women 92%
Similar papers in this journal
- Risk of cancer in regular and low meat-eaters, fish-eaters, and vegetarians: a prospective analysis of UK Biobank participants 93%
- Polygenic Risk Score Improves the Accuracy of a Clinical Risk Score for Coronary Artery Disease 93%
- Separating the effects of early and later life adiposity on colorectal cancer risk: a Mendelian randomization study 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.