Back

A comprehensive genomic framework for identifying genes predisposing to homologous recombination repair deficient breast cancer

Camacho-Valenzuela, J.; Matis, T.; Roca, C.; Cuamatzi Flores, J. L.; Hamel, N.; Rivera, B.; Gravel, S.; Polak, P.; Robles-Espinoza, C. D.; Foulkes, W. D.

2025-10-23 genetic and genomic medicine
10.1101/2025.10.22.25338181 medRxiv
Show abstract

BackgroundPatients with clinical characteristics of increased cancer susceptibility without an identified genetic lesion are regularly seen in clinics. Case-control studies and matched normal/tumour sequencing have advanced the discovery of Cancer Susceptibility Genes (CSGs), with limitations when used independently. We reasoned that combining these strategies alongside mutational signatures and clinical data could improve CSGs identification. MethodsUsing breast cancer exome data from The Cancer Genome Atlas (TCGA-BRCA), we developed a genomic framework that evaluates exome-wide associations of Germline Pathogenic Variants (GPVs) with somatic second hits, within the context of the Homologous Recombination Repair Deficiency (HRD) mutational signature 3 (Sig3). This is complemented by clinico-genomic analysis evaluating clinical and biological plausibility. ResultsOur framework confirmed significant associations with Sig3 of BRCA1/2 GPVs with second hits, validating its performance. THBS4 also reached significance but co-occurred with other HRD-related events. Borderline significance was observed for KIF13B and TESPA1. The clinico-genomics approach further identified KIF13B and TESPA1, as well as RAD51B and other Fanconi Anemia pathway-related genes, which deserve further validation. ConclusionsOur framework strengthens identification of candidate HRD-related breast CSGs through combined statistical and clinico-genomics analyses. It is adaptable to other mutational signatures/cancer types and will benefit from larger and well-annotated datasets.

Matching journals

The top 7 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.