Deep Learning for MYC binding site recognition
Fioresi, r.; Demurtas, P.; Perini, G.
Show abstract
MotivationThe definition of the genome distribution of the Myc transcription factor is extremely important since it may help predict its transcriptional activity particularly in the context of cancer. Myc is among the most powerful oncogenes involved in the occurrence and development of more than 80% of different types of pediatric and adult cancers. Myc regulates thousands of genes which can be in part different, depending on the type of tissues and tumours. Myc distribution along the genome has been determined experimentally through chromatin immunoprecipitation This approach, although powerful, is very time consuming and cannot be routinely applied to tumours of individual patients. Thus, it becomes of paramount importance to develop in-silico tools that can effectively and rapidly predict its distribution on a given cell genome. New advanced computational tools (DeeperBind) can then be successfully employed to determine the function of Myc in a specific tumour, and may help to devise new directions and approaches to experiments first and personalized and more effective therapeutic treatments for a single patient later on. ResultsThe use of DeeperBind with DeepRAM on Colab platform can effectively predict the binding sites for the MYC factor with an accuracy above 0.96 AUC, when trained with multiple cell lines. The analysis of the filters in DeeperBind trained models shows, besides the consensus sequence CACGTG classically associated to the MYC factor, also the other consensus sequences G/C box or TGGGA, respectively bound by the SP1 and MIZ-1 transcription factors, which are known to mediate the MYC repressive response. Overall, our findings suggest a stronger sinergy between the machine learning tools as DeeperBind and biological experiments, which may reduce the time consuming experiments by providing a direction to guide them.
Matching journals
The top 10 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Machine-learning-based Structural Analysis of Interactions between Antibodies and Antigens 91%
- Simulation of multiple microenvironments shows a putative role of RPTPs on the control of Epithelial-to-Mesenchymal Transition. 91%
- miRNAFinder: A Comprehensive Web Resource for Plant Pre-microRNA Classification 91%
Similar papers in this journal
- SubFeat: Feature Subspacing Ensemble Classifier for Function Prediction of DNA, RNA and Protein Sequences 94%
- Optimizing Therapeutic Targets For Breast Cancer Using Boolean Network Models 92%
- Exploring vulnerable building blocks in protein-protein interaction networks of breast tumor and adjacent normal tissues 92%
Similar papers in this journal
- Enhanced performance of gene expression predictive models with protein-mediated spatial chromatin interactions 94%
- A Convolution Based Computational Approach Towards DNA N6-methyladenine Site Identification and Motif Extraction in Rice Genome 94%
- Finding disease modules for cancer and COVID-19 in gene co-expression networks with the Core&Peel method 93%
Similar papers in this journal
- Deepprune: Learning efficient and interpretable convolutional networks through weight pruning for predicting DNA-protein binding 93%
- Tensor decomposition-Based Unsupervised Feature Extraction Applied to Single-Cell Gene Expression Analysis 93%
- Characterization of Human Dosage-Sensitive Transcription Factor Genes 93%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.