Back

Graphical Learning and Causal Inference for Drug Repurposing

Xu, T.; Zhao, J.; Xiomg, M.

2023-08-02 genetic and genomic medicine
10.1101/2023.07.29.23293346 medRxiv
Show abstract

Gene expression profiles that connect drug perturbations, disease gene expression signatures, and clinical data are important for discovering potential drug repurposing indications. However, the current approach to gene expression reversal has several limitations. First, most methods focus on validating the reversal expression of individual genes. Second, there is a lack of causal approaches for identifying drug repurposing candidates. Third, few methods for passing and summarizing information on a graph have been used for drug repurposing analysis, with classical network propagation and gene set enrichment analysis being the most common. Fourth, there is a lack of graph-valued association analysis, with current approaches using real-valued association analysis one gene at a time to reverse abnormal gene expressions to normal gene expressions. To overcome these limitations, we propose a novel causal inference and graph neural network (GNN)-based framework for identifying drug repurposing candidates. We formulated a causal network as a continuous constrained optimization problem and developed a new algorithm for reconstructing large-scale causal networks of up to 1,000 nodes. We conducted large-scale simulations that demonstrated good false positive and false negative rates. To aggregate and summarize information on both nodes and structure from the spatial domain of the causal network, we used directed acyclic graph neural networks (DAGNN). We also developed a new method for graph regression in which both dependent and independent variables are graphs. We used graph regression to measure the degree to which drugs reverse altered gene expressions of disease to normal levels and to select potential drug repurposing candidates. To illustrate the application of our proposed methods for drug repurposing, we applied them to phase I and II L1000 connectivity map perturbational profiles from the Broad Institute LINCS, which consist of gene-expression profiles for thousands of perturbagens at a variety of time points, doses, and cell lines, as well as disease gene expression data under-expressed and over-expressed in response to SARS-CoV-2.

Matching journals

The top 5 journals account for 50% of the predicted probability mass.

1
IEEE/ACM Transactions on Computational Biology and Bioinformatics
38 papers in training set
Top 0.1%
15.5%
2
Scientific Reports
3612 papers in training set
Top 1%
15.3%
3
Heliyon
152 papers in training set
Top 0.1%
10.8%
4
Briefings in Bioinformatics
354 papers in training set
Top 1%
6.8%
5
Computers in Biology and Medicine
128 papers in training set
Top 0.8%
4.4%
50% of probability mass above
6
Computational and Structural Biotechnology Journal
242 papers in training set
Top 1.0%
4.1%
7
Artificial Intelligence in the Life Sciences
13 papers in training set
Top 0.1%
3.5%
8
PLOS ONE
5266 papers in training set
Top 37%
3.3%
9
Frontiers in Molecular Biosciences
102 papers in training set
Top 0.4%
2.4%
10
International Journal of Molecular Sciences
494 papers in training set
Top 5%
2.4%
11
Frontiers in Bioinformatics
49 papers in training set
Top 0.2%
2.4%
12
Frontiers in Genetics
230 papers in training set
Top 2%
2.2%
13
iScience
1154 papers in training set
Top 12%
2.2%
14
Journal of Personalized Medicine
28 papers in training set
Top 0.4%
1.8%
15
Frontiers in Pharmacology
111 papers in training set
Top 2%
1.7%
16
Frontiers in Oncology
103 papers in training set
Top 2%
1.5%
17
Informatics in Medicine Unlocked
22 papers in training set
Top 0.7%
1.4%
18
BMC Genomics
406 papers in training set
Top 6%
1.1%
19
Journal of Translational Medicine
57 papers in training set
Top 2%
0.9%
20
Brain and Behavior
43 papers in training set
Top 2%
0.9%
21
Frontiers in Psychiatry
87 papers in training set
Top 2%
0.9%
22
Patterns
78 papers in training set
Top 3%
0.6%
23
Frontiers in Bioengineering and Biotechnology
98 papers in training set
Top 3%
0.6%
24
Computational Biology and Chemistry
28 papers in training set
Top 1%
0.6%
25
Chaos: An Interdisciplinary Journal of Nonlinear Science
17 papers in training set
Top 0.4%
0.6%