Candidate gene prioritization using graph embedding
DO, Q.; LARMANDE, P.
Show abstract
Candidate genes prioritization allows to rank among a large number of genes, those that are strongly associated with a phenotype or a disease. Due to the important amount of data that needs to be integrate and analyse, gene-to-phenotype association is still a challenging task. In this paper, we evaluated a knowledge graph approach combined with embedding methods to overcome these challenges. We first introduced a dataset of rice genes created from several open-access databases. Then, we used the Translating Embedding model and Convolution Knowledge Base model, to vectorize gene information. Finally, we evaluated the results using link prediction performance and vectors representation using some unsupervised learning techniques.
Matching journals
The top 8 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Cardiac disease diagnosis based on GAN in case of missing data 96%
- Predicting Adverse Drug Effects: A Heterogeneous Graph Convolution Network with a Multi-layer Perceptron Approach 96%
- Regional medical inter-institutional cooperation in medical provider network constructed using patient claims data from Japan 95%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.