Genetic and Cellular Architecture of Breast Cancer Risk in Multi-Ancestry Studies of 159,297 Cases and 212,102 Controls
Li, J. L.; Zanti, M.; Williams, J.; Jahagirdar, O.; Jia, G.; Turcan, A.; Hu, Q.; Brandenburg, J.-T.; Yan, L.; Ho, W.-K.; Li, J.; Miranda, J. P.; Godbole, D.; Dias, J.-A.; Zhang, X.; Dorling, L.; Chen, W. C.; Boddicker, N.; Wang, Y.; Martin, A.; Zhang, Y. D.; Dennis, J.; John, E. M.; Torres-Mejia, G.; Kushi, L.; Weitzel, J.; Neuhausen, S. L.; Carvajal-Carmona, L.; Haiman, C.; Ziv, E.; Fejerman, L.; Zheng, W.; Huo, D.; Easton, D.; Chanock, S. J.; Chatterjee, N.; Kraft, P.; Garcia-Closas, M.; Wong, W. S. W.; Michailidou, K.; Zhu, Q.; Zhang, M. J.; Dutta, D.; Ahearn, T. U.; Zhang, H.
Show abstract
Breast cancer genome-wide association studies (GWAS) have identified over 200 independent genome-wide significant susceptibility markers. However, most studies have focused on one or two ancestral groups. We examined breast cancer genetic architecture using GWAS summary statistics from African (AFR), East Asian (EAS), European (EUR) and Hispanic/Latina (H/L) samples, totaling 159,297 cases and 212,102 controls, comprising the largest multi-ancestry study of breast cancer to date. The logit-scale heritability of breast cancer ranged from h2=0.47 (SE = 0.07) in EAS to AFR h2=0.61 (SE = 0.10), with no significant differences across ancestries (p=0.63). The estimated number of susceptibility markers in a sparse normal-mixture effects model also varied from 4,446 (SE = 3,100) in EAS to 8,308 (SE = 2,751) in AFR, but differences were not significant across ancestries (p=0.55). Cross-sample genetic correlations varied, with the strongest correlation between EUR and EAS ({rho} = 0.79, SE = 0.08) and weakest between AFR and H/L ({rho} = 0.26, SE = 0.24). Common variants in regulatory elements were enriched for genetic association across samples. By integrating the GWAS summary statistics with the Tabula Sapiens scRNA-seq atlas, we identified ancestry-shared associations between breast cancer and specific cell types, including innate immune cells, secretory epithelial cells and stromal cells. Collectively, these results support a largely shared polygenic architecture of breast cancer across ancestries, with consistent enrichment of common regulatory variants and convergent cellular signatures identified through single-cell analyses.
Matching journals
The top 1 journal accounts for 50% of the predicted probability mass.
Similar papers in this journal
- Genome-wide association study identifies 32 novel breast cancer susceptibility loci from overall and subtype-specific analyses 97%
- Leveraging functional genomic annotations and genome coverage to improve polygenic prediction of complex traits within and between ancestries 96%
- Combining case-control status and family history of disease increases association power 96%
Similar papers in this journal
- Evaluating Multi-Ancestry Genome-Wide Association Methods: Statistical Power, Population Structure, and Practical Implications 96%
- Enrichment analyses identify shared associations for 25 quantitative traits in over 600,000 individuals from seven diverse ancestries 96%
- Characterizing substructure via mixture modeling in large-scale genetic summary statistics 96%
Similar papers in this journal
- Analysis across Taiwan Biobank, Biobank Japan and UK Biobank identifies hundreds of novel loci for 36 quantitative traits 96%
- Polygenic scores capture genetic modification of the adiposity-cardiometabolic risk factor relationship 96%
- Polymorphic short tandem repeats make widespread contributions to blood and serum traits 96%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.