Back

DrugPlayGround: Benchmarking Large Language Models and Embeddings for Drug Discovery

Liu, T.; Jiang, S.; Zhang, F.; Sun, K.; Head-Gordon, T.; Zhao, H.

2026-04-07 bioinformatics
10.64898/2026.04.04.716470 bioRxiv
Show abstract

Large language models (LLMs) are in the ascendancy for research in drug discovery, offering unprecedented opportunities to reshape drug research by accelerating hypothesis generation, optimizing candidate prioritization, and enabling more scalable and cost-effective drug discovery pipelines. However there is currently a lack of objective assessments of LLM performance to ascertain their advantages and limitations over traditional drug discovery platforms. To tackle this emergent problem, we have developed DrugPlayGround, a framework to evaluate and benchmark LLM performance for generating meaningful text-based descriptions of physiochemical drug characteristics, drug synergism, drug-protein interactions, and the physiological response to perturbations introduced by drug molecules. Moreover, DrugPlayGround is designed to work with domain experts to provide detailed explanations for justifying the predictions of LLMs, thereby testing LLMs for chemical and biological reasoning capabilities to push their greater use at the frontier of drug discovery at all of its stages.

Matching journals

The top 5 journals account for 50% of the predicted probability mass.

1
Journal of Chemical Information and Modeling
207 papers in training set
Top 0.3%
18.6%
2
Journal of Cheminformatics
25 papers in training set
Top 0.1%
12.5%
3
Bioinformatics
1061 papers in training set
Top 3%
8.4%
4
Briefings in Bioinformatics
326 papers in training set
Top 0.7%
6.8%
5
Computational and Structural Biotechnology Journal
216 papers in training set
Top 0.5%
6.4%
50% of probability mass above
6
Bioinformatics Advances
184 papers in training set
Top 0.7%
4.9%
7
Nucleic Acids Research
1128 papers in training set
Top 6%
3.6%
8
BMC Bioinformatics
383 papers in training set
Top 3%
2.7%
9
npj Systems Biology and Applications
99 papers in training set
Top 0.8%
2.1%
10
PLOS ONE
4510 papers in training set
Top 48%
2.1%
11
Scientific Reports
3102 papers in training set
Top 53%
1.9%
12
Nature Machine Intelligence
61 papers in training set
Top 2%
1.8%
13
Computers in Biology and Medicine
120 papers in training set
Top 2%
1.7%
14
Artificial Intelligence in the Life Sciences
11 papers in training set
Top 0.1%
1.5%
15
PLOS Computational Biology
1633 papers in training set
Top 18%
1.5%
16
Patterns
70 papers in training set
Top 1%
1.2%
17
Metabolites
50 papers in training set
Top 0.7%
1.2%
18
iScience
1063 papers in training set
Top 23%
1.1%
19
International Journal of Molecular Sciences
453 papers in training set
Top 11%
1.1%
20
Advanced Science
249 papers in training set
Top 15%
1.0%
21
Frontiers in Molecular Biosciences
100 papers in training set
Top 4%
0.9%
22
Nature Communications
4913 papers in training set
Top 61%
0.8%
23
npj Digital Medicine
97 papers in training set
Top 3%
0.7%
24
GigaScience
172 papers in training set
Top 3%
0.7%
25
NAR Genomics and Bioinformatics
214 papers in training set
Top 4%
0.7%
26
Chemical Science
71 papers in training set
Top 2%
0.7%
27
Cell Systems
167 papers in training set
Top 14%
0.6%
28
Genome Medicine
154 papers in training set
Top 9%
0.6%