An Entropy-based Directed Random Walk for Pathway Activity Inference Using Topological Importance and Gene Interactions
Tay, H. X.; Sutikno, T.; Kasim, S.; Kasim, S. F. M.; Halim, S. A.; Hassan, R.; Sen, S. C.
Show abstract
The integration of microarray technologies and machine learning methods has become popular in predicting pathological condition of diseases and discovering risk genes. The traditional microarray analysis considers pathways as simple gene sets, treating all genes in the pathway identically while ignoring the pathway networks structure information. This study, however, proposed an entropy-based directed random walk (e-DRW) method to infer pathway activity. This study aims (1) To enhance the gene-weighting method in Directed Random Walk (DRW) by incorporating t-test statistic scores and correlation coefficient values, (2) To implement entropy as a parameter variable for random walking in a biological network, and (3) To apply Entropy Weight Method (EWM) in DRW pathway activity inference. To test the objectives, the gene expression dataset was used as input datasets while the pathway dataset was used as reference datasets to build a directed graph. An equation was proposed to assess the connectivity of nodes in the directed graph via probability values calculated from the Shannon entropy formula. A direct proof of calculation based on the proposed mathematical formula was presented using e-DRW with gene expression data. Based on the results, there was an improvement in terms of sensitivity of prediction and accuracy of cancer classification between e-DRW and conventional DRW. The within-dataset experiments indicated that our novel method demonstrated robust and superior performance in terms of accuracy and number of predicted risk-active pathways compared to the other DRW methods. In conclusion, the results revealed that e-DRW not only improved prediction performance, but also effectively extracted topologically important pathways and genes that are specifically related to the corresponding cancer types.
Matching journals
The top 8 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Improving prediction of drug-target interactions based on fusing multiple features with data balancing and feature selection techniques 97%
- Predicting Adverse Drug Effects: A Heterogeneous Graph Convolution Network with a Multi-layer Perceptron Approach 96%
- Targeting Oncogenic Mutations in Colorectal Cancer using Cryptotanshinone 96%
Similar papers in this journal
- Classification models for Invasive Ductal Carcinoma Progression, based on gene expression data-trained supervised machine learning 97%
- Comparing protein-protein interaction networks of SARS-CoV-2 and (H1N1) influenza using topological features 97%
- A Convolution Based Computational Approach Towards DNA N6-methyladenine Site Identification and Motif Extraction in Rice Genome 96%
Similar papers in this journal
- Investigate the relevance of major signaling pathways in cancer survival using a biologically meaningful deep learning model 95%
- StrongestPath: a Cytoscape application for protein-protein interaction analysis 94%
- Combining Gene Ontology with Deep Neural Networks to Enhance the Clustering of Single Cell RNA-Seq Data 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.