MutSigCVsyn: Identification of Thirty Synonymous Cancer Drivers
Rao, Y.; Ahmed, N.; Pritchard, J.; O'Brien, E.
Show abstract
Synonymous mutations, which change only the DNA sequence but not the encoded protein sequence, can affect protein structure and function, mRNA maturation, and mRNA half-lives. The possibility that synonymous mutations can act as cancer drivers has been explored in several recent studies. However, none of these studies control for all three levels (patient, histology, and gene) of mutational heterogeneity that are known to affect the accurate identification of non-synonymous cancer drivers. Here, we create an algorithm, MutSigCVsyn, an adaptation of MutSigCV, to identify synonymous cancer drivers based on a novel non-coding background model that takes into account the mutational heterogeneity across these levels. Examining 2,572 PCAWG cancer whole-genome sequences, MutSigCVsyn identifies 30 novel synonymous drivers that include mutations in promising candidates like BCL-2. By bringing the best practices in non-synonymous driver identification to the analysis of synonymous drivers, these are promising candidates for future experimental study.
Matching journals
The top 8 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Identification of Relevant Genetic Alterations in Cancer using Topological Data Analysis 96%
- Whole-genome analysis of Nigerian patients with breast cancer reveals ethnic-driven somatic evolution and distinct genomic subtypes 96%
- Machine learning-based tissue of origin classification for cancer of unknown primary diagnostics using genome-wide mutation features 96%
Similar papers in this journal
Similar papers in this journal
- Interpretable deep learning for chromatin-informed inference of transcriptional programs driven by somatic alterations across cancers 96%
- HYENA detects oncogenes activated by distal enhancers in cancer 96%
- NetActivity enhances transcriptional signals by combining gene expression into robust gene set activity scores through interpretable autoencoders 95%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.