Back

ComboPath: An ML system for predicting drug combination effects with superior model specification

Ranasinghe, D. S.; Sanders, N. E.; Tam, H. H.; Liu, C.; Spitz, D.

2024-01-20 biophysics
10.1101/2024.01.16.575408 bioRxiv
Show abstract

Drug combinations have been shown to be an effective strategy for cancer therapy, but identifying beneficial combinations through experiments is labor-intensive and expensive [Mokhtari et al., 2017]. Machine learning (ML) systems that can propose novel and effective drug combinations have the potential to dramatically improve the efficiency of combinatoric drug design. However, the biophysical parameters of drug combinations are degenerate, making it difficult to identify the ground truth of drug interactions even given experimental data of the highest quality available. Existing ML models are highly underspecified to meet this challenge, leaving them vulnerable to producing parameters that are not biophysically realistic and harming generalization. We have developed a new ML model, "ComboPath", aimed at a novel ML task: to predict interpretable cellular dose response surface of a two-drug combination based on each drugs interactions with their known protein targets. ComboPath incorporates a biophysically-motivated intermediate parameterization with prior information used to improve model specification. This is the first ML model to nominate beneficial drug combinations while simultaneously reconstructing the dose response surface, providing insight on both the potential of a drug combination and its optimal dosing for therapeutic development. We show that our models were able to accurately reconstruct 2D dose response surfaces across held out combination samples from the largest available combinatoric screening dataset while substantially improving model specification for key biophysical parameters.

Matching journals

The top 2 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.