Performance of general-population breast cancer risk prediction models in an international consortium
Brantley, K. D.; Ahearn, T. U.; Norton, E. L.; MacInnis, R.; Palmer, J. R.; Fortner, R. T.; Vachon, C. M.; Beane-Freeman, L.; Berrington de Gonzalez, A.; Frost, R.; Bertrand, K. A.; Zirpoli, G.; Neuhouser, M. L.; Barnett, M.; Teras, L. R.; Hodge, J. M.; Patel, A. V.; Bodelon, C.; Lacey, J. V.; Spielfogel, E. S.; Rohan, T. E.; Kirsh, V. A.; Langseth, H.; Tsuruda, K. M.; Milne, R. L.; Haiman, C.; Scott, C. G.; Eliassen, A. H.; Rosner, B.; Willett, W. C.; Romanos-Nanclares, A.; Chen, Y.; Wu, F.; Zheng, W.; Long, J.; O'Brien, K. M.; Sandler, D. P.; Kitahara, C. M.; Linet, M. S.; Anderson, G.; Lars
Show abstract
Background: Several breast cancer (BC) risk prediction models have been developed to provide personal risk assessments. Though individually validated, their performance has not been systematically evaluated across a wide range of populations or ages. Methods: We harmonized individual-level baseline questionnaire data and incident BC diagnoses from 21 cohorts from North America, Europe, and Australia participating in the Breast Cancer Risk Prediction Project. Five-year absolute risk of invasive BC was estimated for five established risk prediction models using classical risk factors only. Discrimination was evaluated by area under the curve (AUC). Calibration was assessed using average and risk-decile specific expected to observed (E/O) ratios. Performance metrics were meta-analyzed across cohorts and models. Metaregression tested associations between cohort characteristics and performance metrics. Results: This analysis included 1,595,977 women aged 20-75 years, enrolled in studies between 1976-2015, with 19,062 (1.2%) invasive BC cases ascertained within 5 years from exposure assessment. Age-adjusted AUCs were similar across models and cohorts (pooled AUCs by model: 0.57-0.58), while E/O ratios varied substantially (pooled E/O ratios by model: 0.83-1.25). Overestimation was common among predicted high-risk individuals (>3%). No appreciable differences in model performance by cohort age, birth year, race, and variable missingness emerged. Calibration improved after assigning race-specific incidence rates. Conclusion: Existing BC risk prediction models provided similar risk discrimination across multiple cohorts, although there was overestimation of risk for high-risk individuals. Performance variation across cohorts was not driven by specific characteristics, which supports development of a unified risk model for diverse populations that leverages appropriate incidence rates.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Novel mammogram-based measures improve breast cancer risk prediction beyond an established mammographic density measure 95%
- Reproductive Factors and Risk of Breast Cancer by Tumor Subtypes among Ghanaian Women: A Population-based Case-control Study 94%
- Long-Term Impact of Stressful Life Events on Breast Cancer Risk: A 36-Year Genetically Informed Prospective Study in the Finnish Twin Cohort 92%
Similar papers in this journal
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.