Single-cell differential expression analysis between conditions within nested settings
Hafner, L.; Sturm, G.; List, M.
Show abstract
Differential expression analysis provides insights into fundamental biological processes and with the advent of single-cell transcriptomics, gene expression can now be studied at the level of individual cells. Many analyses treat cells as samples and assume statistical independence. As cells are pseudoreplicates, this assumption does not hold, leading to reduced robustness, reproducibility, and an inflated type 1 error rate. In this study, we investigate various methods for differential expression analysis on single-cell data, conduct extensive benchmarking and give recommendations for method choice. The tested methods include DESeq2, MAST, DREAM, scVI, the Permutation Test and distinct. We additionally adapt Hierarchical Bootstrapping to differential expression analysis on single-cell data and include it in our benchmark. We found that differential expression analysis methods designed specifically for single-cell data do not offer performance advantages over conventional pseudobulk methods such as DESeq2 when applied to individual data sets. In addition, they mostly require significantly longer run times. For atlas-level analysis, permutation-based methods excel in performance but show poor runtime, suggesting to use DREAM as a compromise between quality and runtime. Overall, our study offers the community a valuable benchmark of methods across diverse scenarios and offers guidelines on method selection.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Heterogeneous pseudobulk simulation enables realistic benchmarking of cell-type deconvolution methods 96%
- omnideconv: a unifying framework for using and benchmarking single-cell-informed deconvolution of bulk RNA-seq data 96%
- A comparison of marker gene selection methods for single-cell RNA sequencing data 96%
Similar papers in this journal
- Optimal tuning of weighted kNN- and diffusion-based methods for denoising single cell genomics data 96%
- Building, Benchmarking, and Exploring Perturbative Maps of Transcriptional and Morphological Data 95%
- Reconstruction Set Test (RESET): a computationally efficient method for single sample gene set testing based on randomized reduced rank reconstruction error 95%
Similar papers in this journal
- A comprehensive comparison on cell type composition inference for spatial transcriptomics data 95%
- scDeepInsight: a supervised cell-type identification method for scRNA-seq data with deep learning 95%
- Sincast: a computational framework to predict cell identities in single cell transcriptomes using bulk atlases as references 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.