OTRec: Deep learning recommender for prospective druggable disease-target associations
Ofer, D.; Linial, M.
Show abstract
Identifying druggable disease-target associations remains a central challenge in translational medicine, limiting therapeutic discovery and repurposing. Here, we present OTRec, a deep learning-based recommender system that ranks such associations at scale and evaluates them in a temporal hold-out setting. Unlike approaches that rely on manually curated or aggregated evidence scores, OTRec employs a two-tower architecture to learn latent representations from 663,351 disease-target pairs. The model integrates heterogeneous inputs, including textual descriptions, ontology-derived features, and biological annotations such as tractability, Gene Ontology (GO) terms, and pathway information. We perform temporal validation by training on the 2022 Open Targets (OT) release and evaluating on clinical trial data from 2025. OTRec improves on the retrospective OT association score (ROC-AUC: 0.872 {+/-} 0.005 vs. 0.559; PR-AUC: 0.288 {+/-}0.009 vs. 0.08). In 5 x 5 target-disjoint cross-validation, OTRec reaches ROC-AUC 0.950 and PR-AUC 0.844) improving on the OT evidence score (ROC-AUC 0.91; PR-AUC 0.45). We rank the druggable genome across [~]19,000 OT platform (OTP) diseases and release [~] 282,500 candidate associations above a 0.65 score threshold (in-distribution CV precision 0.92), covering 4,346 diseases including 2,322 orphan diseases, through an interactive prediction platform. Anonymized demo, code, models, and predictions available: https://anonymous.4open.science/r/OTRec-D6DF/
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- KG-Bench: Benchmarking Graph Neural Network Algorithms for Drug Repurposing 95%
- TUGDA: Task uncertainty guided domain adaptation for robust generalization of cancer drug response prediction from in vitro to in vivo settings 94%
- DrDimont: Explainable drug response predictionfrom differential analysis of multi-omics networks 94%
Similar papers in this journal
- Deep representation learning for clustering longitudinal survival data from electronic health records 95%
- Community assessment of cancer drug combination screens identifies strategies for synergy prediction 94%
- PreMode predicts mode-of-action of missense variants by deep graph representation learning of protein sequence and structural context 93%
Similar papers in this journal
Similar papers in this journal
- Few shot learning for phenotype-driven diagnosis of patients with rare genetic diseases 95%
- Clinical Knowledge Extraction via Sparse Embedding Regression (KESER) with Multi-Center Large Scale Electronic Health Record Data 94%
- Federated Target Trial Emulation using Distributed Observational Data for Treatment Effect Estimation 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.