Comparative fine-mapping of breast cancer susceptibility loci using summary statistics methods and multinomial regression
O'Mahony, D. G.; Beasley, J.; Zanti, M.; Dennis, J.; Dutta, D.; Kraft, P.; Kristensen, V.; Chenevix-Trench, G.; Easton, D. F.; Michailidou, K.
Show abstract
Summary statistics fine-mapping methods offer advantages over classical methods, including avoiding data-sharing constraints and improved modelling of correlated variables and sparse effects. However, its performance has not been comprehensively evaluated in breast cancer using real-world data. Previous multinomial stepwise regression (MNR) fine-mapping analyses for breast cancer identified 196 credible sets. Here, we apply summary statistics fine-mapping, compare methods, and assess parameters influencing performance. Using summary statistics from the Breast Cancer Association Consortium, we compared finiMOM, SuSiE, and FINEMAP to published MNR results across 129 regions. Performance was assessed by recall using in-sample and out-of-sample LD. Discordant credible sets were examined for technical factors, and target genes were defined using the INQUISIT pipeline. SuSiE showed the closest agreement with MNR. Results varied across regions depending on the assumed number of causal variants (L), with higher values reducing recall and no single L maximising performance. At optimal L per region, SuSiE identified 8,192 CCVs in 244 credible sets, with recall of 88%, 86%, and 72% for overall, ER-positive, and ER-negative breast cancer. Thirty MNR sets were missed. Discordance was partially explained by allele flips, imputation quality, and array heterogeneity. Fifty-two MNR-identified genes, including BRCA2, WNT7B and CREBBP were not recovered, while additional candidate genes were identified. Using out-of-sample LD reduced recall by 3% but identified novel variants. Fine-mapping results vary across methods, and no single approach is sufficient. The choice of L strongly influences results, and combining analytical approaches with functional validation can improve causal variant identification.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Cancer PRSweb - an Online Repository with Polygenic Risk Scores (PRS) for Major Cancer Traits and Their Phenome-wide Exploration in Two Independent Biobanks 95%
- A joint transcriptome-wide association study across multiple tissues identifies new candidate susceptibility genes for breast cancer 94%
- Segregation analysis of 17,425 population-based breast cancer families: evidence for genetic susceptibility and risk prediction 93%
Similar papers in this journal
- The mediating role of mammographic density in the protective effect of early-life adiposity on breast cancer risk: a multivariable Mendelian randomization study 93%
- Pan-cancer landscape of homologous recombination deficiency 93%
- Machine learning-based tissue of origin classification for cancer of unknown primary diagnostics using genome-wide mutation features 93%
Similar papers in this journal
- Fine-mapping of 150 breast cancer risk regions identifies 178 high confidence target genes 96%
- Genome-wide association study identifies 32 novel breast cancer susceptibility loci from overall and subtype-specific analyses 94%
- Genomic profiling defines variable clonal relatedness between invasive breast cancer and primary ductal carcinoma in situ 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.