Identification and Characterization of Rare Genetic Variants Associated with pathogenicity Breast Cancer Susceptibility Across Thousand Genome and GnomAD population
Yadav, K.; Prasad, A.
Show abstract
BackgroundInherited variants in cancer susceptibility genes play a critical role in breast cancer risk, yet comprehensive comparative analyses across clinically relevant genes remain limited. While BRCA1 and BRCA2 are well-established high-risk contributors, the broader landscape of germline pathogenicity across breast cancer associated genes warrants systematic evaluation. MethodsWe curated a panel of 15 breast cancer associated genes (BRCA1, BRCA2, TP53, PTEN, CHEK2, ATM, CDH1, PALB2, LSP1, MAP3K1, NF1, RECQL, TOX3, FANCD2, RAD51C) based on clinical and epidemiological relevance. Publicly available germline variant data from gnomAD v4.1 and the 1000 Genomes Project were integrated with ClinVar annotations. Variants were classified by clinical significance and filtered using allele frequency and phenotype annotation to stratify into high-risk and non-high-risk groups. We compared the distribution of variants and assessed penetrance patterns, and we evaluated the discriminatory power of computational pathogenicity predictors including CADD, REVEL, SIFT, and PolyPhen-2. ResultsBRCA1 and BRCA2 harbored the highest burden of pathogenic variants, consistent with their high penetrance status. Moderate-risk genes (PALB2, RAD51C, CHEK2) showed intermediate levels of pathogenicity, while low-risk genes such as RECQL, TOX3, and FANCD2 were predominantly enriched for variants of uncertain significance. High-risk variants exhibited significantly elevated CADD (>30), REVEL (>0.75), and SIFT (<0.05) scores compared to non-high-risk variants (p < 1x102), whereas PolyPhen-2 showed limited discriminatory power. Odds ratio analysis further supported CADD, REVEL, and ClinVar pathogenicity labels as strong discriminators between variant categories. ConclusionsOur analysis reaffirms the pathogenic burden concentrated in high penetrance genes and highlights the utility of combined in silico prediction tools and population frequency filters in variant prioritization. These findings support refined variant curation pipelines for hereditary breast cancer and provide a framework for interpreting gene-level pathogenicity across diverse susceptibility loci.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Estimating diagnostic noise in panel-based genomic analysis 94%
- Classification of Variants of Reduced Penetrance in High Penetrance Cancer Susceptibility Genes: Framework for Genetics Clinicians and Clinical Scientists by CanVIG-UK (Cancer Variant Interpretation Group-UK) 92%
- Evaluation of Bayesian Classification Framework on the Variant Classification of Hereditary Cancer Predisposition Genes 91%
Similar papers in this journal
- Family History Assessment Significantly Enhances Delivery of Precision Medicine in the Genomics Era 93%
- The genetic analysis of a founder Northern American population of European descent identifies FANCI as a candidate familial ovarian cancer risk gene 91%
- Profiling diverse sequence tandem repeats in colorectal cancer reveals co-occurrence of microsatellite and chromosomal instability involving Chromosome 8 91%
Similar papers in this journal
- Characteristics predicting reduced penetrance variants in the high-risk cancer predisposition gene TP53 94%
- Pleiotropy-guided transcriptome imputation from normal and tumor tissues identifies new candidate susceptibility genes for breast and ovarian cancer 93%
- Inverted genomic regions between reference genome builds in humans impact imputation accuracy and decrease the power of association testing 92%
Similar papers in this journal
Similar papers in this journal
- A genome-wide association study of mammographic texture variation 95%
- Common variants in breast cancer risk loci predispose to distinct tumor subtypes 95%
- Putative breast cancer risk variants from populations of South Asian ancestry are under-represented in public variant classification databases 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.