Back

MDMD: a computational model for predicting drug-related microbes based on the aggregated metapaths from a heterogeneous network

Xing, J.; Zhang, X.; Wang, J.

2023-10-13 bioinformatics
10.1101/2023.10.13.562158 bioRxiv
Show abstract

Clinical studies have shown that microbes are closely related to the occurrence of diseases in the human body. It is beneficial for treating diseases by means of microbes to modulate the activity and toxicity of drugs. Therefore, it is significant in predicting associations between drugs and microbes. Recently, there are several computational models for addressing the issue. However, most of them only focus on drug-related microbes and neglect related diseases, which can lead to insufficient training. Here we introduce a new model (called MDMD) is proposed to predict drug-related microbes based on the Metapaths from a heterogeneous network constructed by using the data of Diseases, Microbes, Drugs, the associations of microbe-disease and disease-drug. The MDMD uses an aggregation of the metapath features that can effectively abundance the embedding of the features for different types of nodes and edges in the heterogeneous networks. Then, the MDMD uses the attention mechanism to mark the importance of the metapath vector for each node type which can improve the quality of feature embedding. Experimental results demonstrate that the MDMD improves accuracy by 1.9% compared with other models. The MDMD is also used to predict the microbes of two drugs Lamivudine and Tenofovir which are the antiretroviral drugs used to treat the Acquired Immune Deficiency Syndrome(AIDS). The results show that 90-95% of microbes are reported in the PubMed. Mycobacterium tuberculosis(Mtb) is a specific microbe only predicted by the MDMD. An online platform of the MDMD is available in https://mdmd2023.bit1024.top/, in which the source code of the MDMD and the data in the work can be downloaded. Author summaryMicrobes inhabit multiple organs of the human body that consist of bacteria, fungi, and viruses. Extensive research shows that the microbes can adjust the efficacy and toxicity of drugs to treat the disease. The efficient and accurate selection of drug-related microbes is important for drug research and disease treatment. However, screening of drug-related microbes relies on traditional lab experiments that are labor-intensive and costly. With the growth of high-throughput data, the research of drug-related microbes urgently needs a computational method in bioinformatics. However, most of them only focus on drug-related microbes and neglect related diseases, which can lead to insufficient training. Therefore, we propose a new method (called MDMD) based on the aggregation of the metapath to efficiently and accurately predict potential drug-related microbes within the microbes-disease-drug network.

Matching journals

The top 7 journals account for 50% of the predicted probability mass.

1
IEEE Transactions on Computational Biology and Bioinformatics
20 papers in training set
Top 0.1%
11.6%
2
Briefings in Bioinformatics
354 papers in training set
Top 0.5%
11.6%
3
IEEE/ACM Transactions on Computational Biology and Bioinformatics
38 papers in training set
Top 0.1%
9.4%
4
Bioinformatics
1204 papers in training set
Top 4%
6.6%
5
PLOS Computational Biology
1863 papers in training set
Top 7%
5.4%
6
Computers in Biology and Medicine
128 papers in training set
Top 0.6%
5.3%
7
IEEE Journal of Biomedical and Health Informatics
37 papers in training set
Top 0.2%
4.2%
50% of probability mass above
8
BMC Bioinformatics
457 papers in training set
Top 2%
3.9%
9
PLOS ONE
5266 papers in training set
Top 35%
3.9%
10
IEEE Access
35 papers in training set
Top 0.4%
3.2%
11
Genomics, Proteomics & Bioinformatics
172 papers in training set
Top 0.7%
3.1%
12
Patterns
78 papers in training set
Top 1%
2.1%
13
Journal of Computational Biology
48 papers in training set
Top 0.5%
2.1%
14
GigaScience
212 papers in training set
Top 2%
2.1%
15
Journal of Biomedical Informatics
47 papers in training set
Top 0.7%
1.9%
16
Computational and Structural Biotechnology Journal
242 papers in training set
Top 4%
1.7%
17
Scientific Reports
3612 papers in training set
Top 56%
1.7%
18
Expert Systems with Applications
11 papers in training set
Top 0.2%
1.6%
19
Bioinformatics Advances
203 papers in training set
Top 3%
1.6%
20
Journal of Medical Internet Research
87 papers in training set
Top 2%
1.6%
21
BMC Medical Informatics and Decision Making
43 papers in training set
Top 1%
1.1%
22
Journal of Chemical Information and Modeling
238 papers in training set
Top 2%
1.0%
23
iScience
1154 papers in training set
Top 37%
0.8%
24
BioData Mining
22 papers in training set
Top 1%
0.6%
25
Life
29 papers in training set
Top 1%
0.6%
26
Frontiers in Genetics
230 papers in training set
Top 7%
0.6%