MERIT: Mechanism driven model predicts drug outcomes and nominates indications for failed drugs
Koh-Tan, H. H. C.; Meic, I.; Sarı, B. A.; Muller, S.; Richman, G.
Show abstract
Drug development depends on efficacy and safety, but many trial-outcome prediction models incorporate trial design, prior development history or compound identity, enabling compound memorization and inflating apparent performance. We developed MEchanism-Resolved Inference of Trial outcomes (MERIT), a model that predicts trial outcomes from molecular and disease features without using information on similar-compound success. MERIT integrates the disease and drug of interest with large-scale drug-protein, protein-metabolite and immune interaction maps to link a drugs intended and potential off-target effects to tissue-specific efficacy and safety. Across 753 small-molecule drugs and 3,133 trials, MERIT achieved a best-in-class overall AUROC of 0.770 (0.765 for efficacy and 0.784 for safety). MERIT also recovered the eventual approved indications for 83% of failed drugs. Finally, we registered locked, outcome-blind predictions for 55 drug-indication pairs in ongoing Phase III trials, establishing a prospective evaluation cohort.
Matching journals
The top 8 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Community assessment of cancer drug combination screens identifies strategies for synergy prediction 96%
- First fully-automated AI/ML virtual screening cascade implemented at a drug discovery centre in Africa 94%
- In silico discovery of nanobody binders to a G-protein coupled receptor using AlphaFold-Multimer 94%
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
- PSICHIC: physicochemical graph neural network for learning protein-ligand interaction fingerprints from sequence data 94%
- TrustAffinity: accurate, reliable and scalable out-of-distribution protein-ligand binding affinity prediction using trustworthy deep learning 94%
- Sagittarius: Extrapolating Heterogeneous Time-Series Gene Expression Data 93%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.