A Systematic Comparison of Single-Cell Perturbation Response Prediction Models
Li, L.; You, Y.; Liao, W.; Fan, X.; Lu, S.; Cao, Y.; Li, B.; Ren, W.; Fu, Y.; Kong, J.; Zheng, S.; Chen, J.; Liu, X.; Tian, L.
Show abstract
Predicting single-cell transcriptional responses to perturbations is central to dissecting gene regulation and accelerating therapeutic design, yet the field lacks a rigorous, task-spanning assessment of model behavior. We present a large-scale benchmark of 12 representative methods and 3 baselines across 25 datasets spanning diverse perturbation modalities and species, including two new primary immune-cell drug-response resources. We evaluated three core tasks--generalization to unseen single-gene perturbations, prediction of combinatorial interactions, and transfer across cell types--using 24 metrics covering expression-level accuracy, relative changes, differential expression recovery, and distributional similarity. Across tasks, performance depended strongly on perturbation effect size and evaluation perspective: expression-level agreement was highest for small-effect perturbations resembling controls, whereas delta- and DE-based metrics improved with larger effects, providing clearer signals. Models shared a conservative bias, with fine-tuned foundation models compressing variance and underestimating synergistic effects in combinations. PerturbNet showed superior recovery of DE signatures in Tasks 1 and 2, while no method consistently generalized across cell types in Task 3, where dataset heterogeneity dominated outcomes. This benchmark establishes current methodological limits, clarifies when metrics diverge, and provides a foundation for developing virtual-cell models that more faithfully capture heterogeneous perturbation responses.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- CellFM: a large-scale foundation model pre-trained on transcriptomics of 100 million human cells 97%
- Triple-effect correction for Cell Painting data with contrastive and domain-adversarial learning 96%
- Normalisr: normalization and association testing for single-cell CRISPR screen and co-expression 96%
Similar papers in this journal
- Capturing cell heterogeneity in representations of cell populations for image-based profiling using contrastive learning 97%
- Building, Benchmarking, and Exploring Perturbative Maps of Transcriptional and Morphological Data 96%
- Non-linear Archetypal Analysis of Single-cell RNA-seq Data by Deep Autoencoders 96%
Similar papers in this journal
- Integrating temporal single-cell gene expression modalities for trajectory inference and disease prediction 96%
- scDesign2: a transparent simulator that generates high-fidelity single-cell gene expression count data with gene correlations captured 96%
- scAlign: a tool for alignment, integration and rare cell identification from scRNA-seq data 96%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.