CREWdb: Optimizing Chromatin Readers, Erasers, and Writers Database using Machine Learning-Based Approach
Natesan, M.; Kong, M.; Shi, M.; Ghag, R.; Mollah, S. A.
Show abstract
Aberration in heterochromatin and euchromatin states contributes to various disease phenotypes. The transcriptional regulation between these two states is significantly governed by post-translational modifications made by three functional types of chromatin regulators: readers, writers, and erasers. Writers introduce a chemical modification to DNA and histone tails, readers bind the modification to histone tails using specialized domains, and erasers remove the modification introduced by writers. Altered regulation of these chromatin regulators results in complex diseases such as cancer, neurodevelopmental diseases, myocardial diseases, kidney diseases, and embryonic development. Due to the reversible nature of chromatin modifications, we can develop therapeutic approaches targeting these chromatin regulators. However, a limited number of chromatin regulators have been identified thus far, and a subset of them are ambiguously classified as multiple chromatin regulator functional types. Thus, we have developed machine learning-based approaches to predict and classify the functional roles of chromatin regulator proteins, thereby optimizing the accuracy of the first comprehensive database of chromatin regulators known as CREWdb. GitHub URLCREWdb source code is available at https://github.com/smollahlab/CREWdb Database URLCREWdb webtool is available at http://mollahlab.wustl.edu/crewdb
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- iMDA-BN: Identification of miRNA-Disease Associations based on the Biological Network and Graph Embedding Algorithm 94%
- PRIMITI: a computational approach for accurate prediction of miRNA-target mRNA interaction 94%
- Modeling and analysis of site-specific mutations in cancer identifies known plus putative novel hotspots and bias due to contextual sequences 94%
Similar papers in this journal
- Crinet: A computational tool to infer genome-wide competing endogenous RNA (ceRNA) interactions 94%
- All of gene expression (AOE): an integrated index for public gene expression databases 94%
- GeneTEFlow: A Nextflow-based pipeline for analysing gene and transposable elements expression from RNA-Seq data 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.