Back

F.A.D.E. (Fully Agentic Drug Engine): A Conversational AI Platform for Drug Discovery

Kantorow, J.; Mani, N.; Mohanraj, N. R.; Zong, X.

2026-06-25 biophysics
10.64898/2026.06.20.733481 bioRxiv
Show abstract

Drug discovery remains one of the costliest and most time-intensive endeavors in the pharmaceutical pipeline, with average development costs exceeding $2.3 billion per drug, timelines spanning more than a decade, and attrition rates above 90% in clinical trials. While computational methods have expanded the searchable chemical space, current pipelines remain fragmented and largely inaccessible to researchers without deep interdisciplinary expertise. Here we present F.A.D.E. (Fully Agentic Drug Engine), a multi-agent, open-source platform that converts natural language queries into potential drug candidates, substantially lowering the expertise barrier to advanced computational drug discovery. F.A.D.E. employs a three-branch hierarchical architecture that adapts to the level of available structural data for any protein target, integrating structure prediction, binding pocket detection, equivariant diffusion-based de novo ligand generation, and binding affinity estimation into a single automated pipeline. We validate F.A.D.E. on two structurally distinct targets: the epidermal growth factor receptor kinase domain (EGFR), a well-established oncology target, and cellular retinol-binding protein 1 (CRBP1), a lipid-binding protein involved in retinoid metabolism. For EGFR, our generated candidates achieved QED scores of 0.85 compared to 0.46 for the co-crystallised reference ligand, demonstrating marked improvement in predicted drug-likeness. Results across both targets confirm that F.A.D.E. can reliably generate chemically tractable, drug-like hit compounds across diverse protein classes from simple natural language input.

Matching journals

The top 1 journal accounts for 50% of the predicted probability mass.

1
Journal of Chemical Information and Modeling
238 papers in training set
Top 0.1%
50.2%
50% of probability mass above
2
Bioinformatics Advances
203 papers in training set
Top 0.1%
12.7%
3
Bioinformatics
1204 papers in training set
Top 6%
2.6%
4
Nature Methods
385 papers in training set
Top 4%
1.9%
5
Proceedings of the National Academy of Sciences
2444 papers in training set
Top 26%
1.9%
6
Nature Communications
5641 papers in training set
Top 44%
1.9%
7
Journal of Chemical Theory and Computation
140 papers in training set
Top 0.8%
1.7%
8
PLOS Computational Biology
1863 papers in training set
Top 14%
1.7%
9
PLOS ONE
5266 papers in training set
Top 49%
1.7%
10
Scientific Reports
3612 papers in training set
Top 59%
1.5%
11
Computational and Structural Biotechnology Journal
242 papers in training set
Top 4%
1.3%
12
eLife
5828 papers in training set
Top 58%
1.1%
13
Protein Science
246 papers in training set
Top 3%
1.1%
14
Frontiers in Pharmacology
111 papers in training set
Top 2%
1.1%
15
ACS Omega
105 papers in training set
Top 3%
1.0%
16
Chemical Science
73 papers in training set
Top 2%
0.8%
17
Communications Chemistry
48 papers in training set
Top 2%
0.8%
18
Journal of Cheminformatics
29 papers in training set
Top 0.7%
0.8%
19
Nucleic Acids Research
1281 papers in training set
Top 14%
0.8%
20
iScience
1154 papers in training set
Top 36%
0.8%
21
Cell Systems
201 papers in training set
Top 5%
0.8%
22
Journal of Molecular Biology
232 papers in training set
Top 4%
0.8%