Computational structure prediction methods enable the systematic identification of oncogenic mutations
Fu, X.; Reglero, C.; Swamy, V.; Loh, J. W.; Khiabanian, H.; Albero, R.; Forouhar, F.; Al Quraishi, M.; Ferrando, A. A.; Rabadan, R.
Show abstract
Oncogenic mutations are associated with the activation of key pathways necessary for the initiation, progression and treatment-evasion of tumors. While large genomic studies provide the opportunity of identifying these mutations, the vast majority of variants have unclear functional roles presenting a challenge for the use of genomic studies in the clinical/therapeutic setting. Recent developments in predicting protein structures enable the systematic large-scale characterization of structures providing a link from genomic data to functional impact. Here, we observed that most oncogenic mutations tend to occur in protein regions that undergo conformation changes in the presence of the activating mutation or when interacting with a protein partner. By combining evolutionary information and protein structure prediction, we introduce the Evolutionary and Structure (ES) score, a computational approach that enables the systematic identification of hotspot somatic mutations in cancer. The predicted sites tend to occur in Short Linear Motifs and protein-protein interfaces. We test the use of ES-scores in genomic studies in pediatric leukemias that easily recapitulates the main mechanisms of resistance to targeted and chemotherapy drugs. To experimentally test the functional role of the predictions, we performed saturated mutagenesis in NT5C2, a protein commonly mutated in relapsed pediatric lymphocytic leukemias. The approach was able to capture both commonly mutated sites and identify previously uncharacterized functionally relevant regions that are not frequently mutated in these cancers. This work shows that the characterization of protein structures provides a link between large genomic studies, with mostly variants of unknown significance, to functional systematic characterization, prioritizing variants of interest in the therapeutic setting and informing on their possible mechanisms of action.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Identification of Relevant Genetic Alterations in Cancer using Topological Data Analysis 96%
- Single-cell transcriptomes identify patient-tailored therapies for selective co-inhibition of cancer clones 96%
- Joint profiling of DNA and proteins in single cells to dissect genotype-phenotype associations in leukemia 96%
Similar papers in this journal
- Genetic dependencies associated with transcription factor activities in human cancer cell lines 96%
- Interrogation of cancer gene dependencies reveals novel paralog interactions of autosome and sexchromosome encoded genes 96%
- Distinct mutational processes shape selection of MHC class I and class II mutations across primary and metastatic tumors 95%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.