Deep Prediction of Human Essential Genes using Weighted Protein-Protein Interaction Networks
Mehrpou, S.; Mansoori, E.
Show abstract
Essential proteins are group of proteins that are indispensable to survival and development of cells. Prediction and analysis of essential genes/proteins are crucial for uncovering the mechanisms of cells. Using bioinformatics and high-throughput technologies, forecasting essential genes/proteins by protein-protein interaction (PPI) networks have become more efficient than traditional approaches which use expensive and time-consuming experimental methods. Previous studies have found that the essentiality of genes closely relates to their properties in PPI network. In this work, we propose a supervised deep model for predicting human essential genes using neighboring details of genes/proteins in the PPI network. Our approach implements a weight-biased random walk on PPI network to get the node network context. Then, some different measures are used to get some feature vectors for each node (gene/protein) that preserve the network structure as well as the genes properties in the PPI network. These feature vectors are then fed to a Relational AutoEncoder to embed the genes features into latent space. At last, these embedded features are put into a trained classifier to predict the human essential genes. The prediction results on two human PPI networks show that our model achieves better performance than those that only refer to genes centrality properties in the network.
Matching journals
The top 8 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- XGSEA: CROSS-species Gene Set Enrichment Analysis via domain adaptation 97%
- Predicting drug-target interactions using multi-label learning with community detection method (DTI-MLCD) 97%
- MeSHHeading2vec: A new method for representing MeSH headings as feature vectors based on graph embedding algorithm 97%
Similar papers in this journal
- Building explainable graph neural network by sparse learning for the drug-protein binding prediction 95%
- Combined topological data analysis and geometric deep learning reveal niches by the quantification of protein binding pockets 95%
- Sensitivity analysis of genome-scale metabolic flux prediction 94%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.