Integrative Gene-Centric Analysis of Breast Cancer: High-Confidence Germline Predisposition Genes from Population-Scale Cohorts
Zucker, R.; Schreiber, S.; Stern, A.; Linial, M.
Show abstract
BackgroundHeritable breast cancer (BC) predisposition is shaped by high-penetrance genes such as BRCA1 and BRCA2, but many moderate- and low-penetrance genes remain uncharacterized. While over 100 independent loci have been reported, their causal genes are often mixed with spurious identifications. MethodsUsing large-scale, multi-ethnic genomic data from cohorts including the UK Biobank (UKB) and FinnGen (FG), we refined the catalog of BC predisposition genes by requiring consistency across multiple GWAS in Open Targets (OT) and excluding likely false positives, thereby reaffirming the contribution of established genes such as BRCA1, BRCA2, PALB2, CHEK2, and other DNA repair genes. ResultsOur multi-cohort design enabled replication across European ancestry, although transferability to other populations was limited, as shown in the Million Veteran Program (MVP) cohort. By integrating genome, transcriptome-, and proteome-wide association studies (GWAS, TWAS, and PWAS), we identified 38 high-confidence BC predisposition genes, including eight previously reported drivers, 13 genes supported by multiple lines of evidence, and a set of candidates (e.g., APOBEC3A, TNS1, PEX14) that currently lack supporting evidence in the context of heritable BC. PWAS, a gene-based association approach that assess the aggregated effect of coding variants on protein function was applied for the entire UKB population. A dozen candidate genes were identified, most of them were considered within the core gene set, and additional two genes that match recessive inheritance. Overlap with FG data and additional gene-based models further supported the involvement of moderate- and high-penetrance genes such as DNMT3A and ATM. ConclusionsThese findings refine the genetic architecture of BC susceptibility, define a core set of predisposition genes, and highlight novel candidates for functional follow-up. The core list of BC susceptibility genes provides biological insights and a framework for gene-based targeted prevention and clinical risk prediction.
Matching journals
The top 9 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Validation of the BOADICEA Model in a Prospective Cohort of BRCA1/2 Pathogenic Variant Carriers 94%
- A Comprehensive Epithelial Tubo-Ovarian Cancer Risk Prediction Model Incorporating Genetic and Epidemiological Risk Factors 93%
- Heritable genetic variants in key cancer genes link cancer risk with anthropometric traits 93%
Similar papers in this journal
- Identifying therapeutic targets for cancer: 2,094 circulating proteins and risk of nine cancers 95%
- The mediating role of mammographic density in the protective effect of early-life adiposity on breast cancer risk: a multivariable Mendelian randomization study 95%
- Pan-cancer landscape of homologous recombination deficiency 94%
Similar papers in this journal
Similar papers in this journal
- Co-expression in tissue-specific gene networks links genes in cancer-susceptibility loci to known somatic driver genes 94%
- Dynamic clustering of genomics cohorts beyond race, ethnicity--and ancestry 94%
- Clinically relevant combined effect of polygenic background, rare pathogenic germline variants, and family history on colorectal cancer incidence 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.