scID: Identification of equivalent transcriptional cell populations across single cell RNA-seq data using discriminant analysis
Boufea, K.; Seth, S.; Batada, N. N.
Show abstract
The power of single cell RNA sequencing (scRNA-seq) stems from its ability to uncover cell type-dependent phenotypes, which rests on the accuracy of cell type identification. However, resolving cell types within and, thus, comparison of scRNA-seq data across conditions is challenging due to technical factors such as sparsity, low number of cells and batch effect. To address these challenges we developed scID (Single Cell IDentification), which uses the framework of Fishers Linear Discriminant Analysis to identify transcriptionally related cell types between scRNA-seq datasets. We demonstrate the accuracy and performance of scID relative to existing methods on several published datasets. By increasing power to identify transcriptionally similar cell types across datasets, scID enhances investigators ability to extract biological insights from scRNA-seq data.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- cellHarmony: Cell-level matching and holistic comparison of single-cell transcriptomes 96%
- Recovering false negatives in CRISPR fitness screens with JLOE 94%
- Single-cell Genome-and-Transcriptome sequencing without upfront whole-genome amplification reveals cell state plasticity of melanoma subclones 94%
Similar papers in this journal
Similar papers in this journal
- Automated quality control and cell identification of droplet-based single-cell data using dropkick 95%
- Highly accurate reference and method selection for universal cross-dataset cell type annotation with CAMUS 95%
- Dynamic Analysis of Alternative Polyadenylation from Single-Cell RNA-Seq(scDaPars) Reveals Cell Subpopulations Invisible to Gene Expression Analysis 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.