Detecting regions of differential abundance betweenscRNA-Seq datasets
Zhao, J.; Jaffe, A.; Cheng, X.; Flavell, R.; Kluger, Y.
Show abstract
Traditional cell clustering analysis used to compare the transcriptomic landscapes between two biological states in single cell RNA sequencing (scRNA-seq) is largely inadequate to functionally identify distinct and important differentially abundant (DA) subpopulations between groups. This problem is exacerbated further when using unsupervised clustering approaches where differences are not observed in clear cluster structure and therefore many important differences between two biological states go entirely unseen. Here, we develop DA-seq, a powerful unbiased, multi-scale algorithm that uniquely detects and decodes novel DA subpopulations not restricted to well separated clusters or known cell types. We apply DA-seq to several publicly available scRNA-seq datasets on various biological systems to detect differences between distinct phenotype in COVID-19 cases, melanomas subjected to immune checkpoint therapy, embryonic development and aging brain, as well as simulated data. Importantly, we find that DA-seq not only recovers the DA cell types as discovered in the original studies, but also reveals new DA subpopulations that were not described before. Analysis of these novel subpopulations yields new biological insights that would otherwise be neglected.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- eSVD-DE: Cohort-wide differential expression in single-cell RNA-seq data using exponential-family embeddings 97%
- scTensor detects many-to-many cell-cell interactions from single cell RNA-sequencing data 95%
- A Markov Random Field Model for Network-based Differential Expression Analysis of Single-cell RNA-seq Data 95%
Similar papers in this journal
- MOSCATO: A Supervised Approach for Analyzing Multi-Omic Single-Cell Data 95%
- Functional module detection through integration of single-cell RNA sequencing data with protein-protein interaction networks 95%
- Probability of stealth multiplets in sample-multiplexing for droplet-based single-cell analysis 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.