The EFFECT benchmark suite: measuring cancer sensitivity prediction performance - without the bias
Szalai, B.; Gaspar, I.; Kaszas, V.; Mero, L.; Sztilkovics, M.; Szalay, K. Z.
Show abstract
1.Creating computational biology models applicable to industry is much more difficult than it appears. There is a major gap between a model that looks good on paper and a model that performs well in the drug discovery process. We are trying to shrink this gap by introducing the Evaluation Framework For predicting Efficiency of Cancer Treatment (EFFECT) benchmark suite based on the DepMap and GDSC data sets to facilitate the creation of well-applicable machine learning models capable of predicting gene essentiality and/or drug sensitivity on in vitro cancer cell lines. We show that standard evaluation metrics like Pearson correlation are misleading due to inherent biases in the data. Thus, to assess the performance of models properly, we propose the use of cell line/perturbation exclusive data splits, perturbation-wise evaluation, and the application of our Bias Detector framework, which can identify model predictions not explicable by data bias alone. Testing the EFFECT suite on a few popular machine learning (ML) models showed that while library-standard non-linear models have measurable performance in splits representing precision medicine and target identification tasks, the actual corrected correlations are rather low, showing that even simple knock-out (KO)/drug sensitivity prediction is a yet unsolved task. For this reason, we aim our proposed framework to be a unified test and evaluation pipeline for ML models predicting cancer sensitivity data, facilitating unbiased benchmarking to support teams to improve on the state of the art.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Mechanistic model of MAPK signaling reveals how allostery and rewiring contribute to drug resistance 94%
- Drug mechanism-of-action discovery through the integration of pharmacological and CRISPR screens 94%
- Cell-to-cell and type-to-type heterogeneity of signaling networks: Insights from the crowd 94%
Similar papers in this journal
- Predicting drug polypharmacology from cell morphology readouts using variational autoencoder latent space arithmetic 97%
- Capturing cell heterogeneity in representations of cell populations for image-based profiling using contrastive learning 95%
- Causal reasoning over knowledge graphs leveraging drug-perturbed and disease-specific transcriptomic signatures for drug discovery 95%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.