Back

A Bayesian active learning platform for scalable combination drug screens

Tosh, C.; Tec, M.; White, J.; Quinn, J. F.; Ibanez Sanchez, G.; Calder, P.; Kung, A. L.; Dela Cruz, F. S.; Tansey, W.

2023-12-19 bioinformatics
10.1101/2023.12.18.572245 bioRxiv
Show abstract

Large-scale combination drug screens are generally considered intractable due to the immense number of possible combinations. Existing approaches use ad hoc fixed experimental designs then train machine learning models to impute novel combinations. Here we propose BATCHIE, an orthogonal approach that conducts experiments dynamically in batches. BATCHIE uses information theory and probabilistic modeling to design each batch to be maximally informative based on the results of previous experiments. On retrospective experiments from previous large-scale screens, BATCHIE designs rapidly discover highly effective and synergistic combinations. To validate BATCHIE prospectively, we conducted a combination screen on a collection of pediatric cancer cell lines using a 206 drug library. After exploring only 4% of the 1.4M possible experiments, the BATCHIE model was highly accurate at predicting novel combinations and detecting synergies. Further, the model identified a panel of top combinations for Ewing sarcomas, all of which were experimentally confirmed to be effective, including the rational and translatable top hit of PARP plus topoisomerase I inhibition. These results demonstrate that adaptive experiments can enable large-scale unbiased combination drug screens with a relatively small number of experiments, thereby powering a new wave of combination drug discoveries. BATCHIE is open source and publicly available (https://github.com/tansey-lab/batchie).

Matching journals

The top 3 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.