Back

A Bayesian Network-Based Framework for Causal Cancer Drug Target Discovery Integrating Patient and Cell Line Data

Yoon, S. H.; Park, Y. R.; Kim, H. U.

2026-07-13 bioinformatics
10.64898/2026.07.13.736676 bioRxiv
Show abstract

Current approaches to cancer drug target discovery face two key limitations: poor translation of cell line-derived targets to patient tumors, and the lack of causal explanation of the regulatory mechanisms underlying target prioritization. Here we present BayesTx (Bayesian Therapeutics target discovery), a Bayesian network framework that integrates patient transcriptomics data with cell line data to identify causal therapeutic targets in cancer. BayesTx projects both data domains into a shared biological space of pathway and transcription factor activities, learns domain-specific causal graphs, and merges them through weighted edge aggregation with bootstrap consensus filtering. Do-simulation on the consensus network quantifies the causal effect of each transcription factor on cancer cell viability. Applied to breast cancer using TCGA-BRCA (The Cancer Genome Atlas breast cancer cohort) and DepMap (Cancer Dependency Map) datasets, the framework ranked 47 transcription factors by predicted causal impact, with gene-level targets further derived through regulon-based propagation. Top-ranked transcription factor (TF) targets were independently supported by survival analysis in external cohort data and pharmacogenomic drug response associations. Overall, BayesTx demonstrates that cross-domain Bayesian network modeling can bridge patient and cell line data to systematically identify causal therapeutic targets in cancer.

Matching journals

The top 7 journals account for 50% of the predicted probability mass.

1
Bioinformatics
1204 papers in training set
Top 2%
15.0%
2
PLOS Computational Biology
1863 papers in training set
Top 2%
12.6%
3
Nature Communications
5641 papers in training set
Top 24%
6.7%
4
npj Systems Biology and Applications
125 papers in training set
Top 0.2%
6.2%
5
JCO Clinical Cancer Informatics
22 papers in training set
Top 0.2%
4.8%
6
PLOS ONE
5266 papers in training set
Top 33%
4.3%
7
BMC Bioinformatics
457 papers in training set
Top 2%
4.0%
50% of probability mass above
8
Briefings in Bioinformatics
354 papers in training set
Top 2%
4.0%
9
Scientific Reports
3612 papers in training set
Top 34%
3.2%
10
NAR Genomics and Bioinformatics
242 papers in training set
Top 2%
2.4%
11
Cell Systems
201 papers in training set
Top 2%
2.1%
12
Nucleic Acids Research
1281 papers in training set
Top 8%
1.9%
13
Cancer Research
130 papers in training set
Top 2%
1.7%
14
Nature Methods
385 papers in training set
Top 4%
1.7%
15
Proceedings of the National Academy of Sciences
2444 papers in training set
Top 28%
1.7%
16
BioData Mining
22 papers in training set
Top 0.3%
1.7%
17
Patterns
78 papers in training set
Top 1%
1.7%
18
Bioinformatics Advances
203 papers in training set
Top 3%
1.5%
19
Journal of the American Medical Informatics Association
71 papers in training set
Top 2%
1.3%
20
eLife
5828 papers in training set
Top 58%
1.1%
21
Genome Medicine
183 papers in training set
Top 4%
1.1%
22
Biometrics
23 papers in training set
Top 0.3%
1.0%
23
Computational and Structural Biotechnology Journal
242 papers in training set
Top 7%
0.8%
24
Science Advances
1243 papers in training set
Top 30%
0.8%
25
Genome Biology
637 papers in training set
Top 9%
0.8%
26
Frontiers in Artificial Intelligence
20 papers in training set
Top 0.8%
0.8%
27
BMC Genomics
406 papers in training set
Top 8%
0.8%
28
npj Digital Medicine
118 papers in training set
Top 3%
0.8%
29
GigaScience
212 papers in training set
Top 5%
0.8%
30
Nature Machine Intelligence
70 papers in training set
Top 3%
0.6%