A mechanism-annotated benchmark reveals limited fidelity to drug-response signatures in single-cell perturbation models
Li, L.; Duan, S.; Zha, X.; Ye, F.; Zhang, Y.; Zhang, X.; Cao, Y.; Liu, C.
Show abstract
Single-cell drug perturbation models are increasingly used to predict how compounds remodel cellular states, but they are still largely assessed by expression reconstruction. Whether high expression similarity reflects preservation of drug-response signatures remains unclear. Here we present scDrugPerturb-Bench, a mechanism-annotated benchmark that links matched control and drug-treated single-cell RNA-sequencing profiles to literature-curated directional key-gene evidence. The resource covers 181 datasets, 423 annotated response cases, 717 unique key genes and 2.5 million cells. We introduce the Mechanism Fidelity Score (MFS) to evaluate key-gene direction, effect-size recovery, gene-set coherence, mechanism specificity and pathway-level response polarity. Across 12 perturbation-prediction models, 3 baselines and 10 data splits, expression-similarity metrics were weakly aligned with MFS and selected different model configurations. Mechanism-aware selection improved early drug retrieval in a transcriptome-based drug design evaluation, indicating that MFS provides practical information beyond benchmark reporting. Systematic benchmarking revealed limited fidelity to drug-response signatures across cell-line and source-integrated settings. Frozen single-cell foundation model embeddings produced local, metric-dependent gains rather than universal improvements, and source context substantially reshaped model assessment. Hard-negative tests further showed that plausible perturbation responses can arise from non-specific transcriptional shortcuts. These results show that expression reconstruction is an insufficient proxy for preserving drug-response signatures and establish scDrugPerturb-Bench as a benchmark for mechanism-aware evaluation of single-cell drug perturbation models.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- EternaBrain: Automated RNA design through move sets from an Internet-scale RNA videogame 94%
- Impact of between-tissue differences on pan-cancer predictions of drug sensitivity 92%
- Capturing cell heterogeneity in representations of cell populations for image-based profiling using contrastive learning 92%
Similar papers in this journal
- Predicting anti-cancer drug synergy using extended drug similarity profiles 94%
- Hybrid Deep Learning with Protein Language Models and Dual-Path Architecture for Predicting IDP Functions 91%
- Learning interpretable cellular embedding for inferring biological mechanisms underlying single-cell transcriptomics 91%
Similar papers in this journal
- Signatures of cell death and proliferation in perturbation transcriptomics data - from confounding factor to effective prediction 94%
- LCA robustly reveals subtle diversity in large-scale single-cell RNA-seq data 92%
- Single-Cell Trajectory Inference for Detecting Transient Events in Biological Processes 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.