UnBlender: validating individual analyses in respiratory bulk RNA-seq cell type deconvolution
Gillett, T. E.; van den Berge, M.; Nawijn, M. C.; Koppelman, G. H.
Show abstract
Analysis of RNA-seq data of respiratory samples has contributed much to our understanding of lung disease. However, bulk RNA-seq data are dependent on both cell type composition and the transcriptional activity of these samples constituent cells, which complicates interpretation. Cell type deconvolution is frequently used to estimate cell type proportions of bulk transcriptomic gene expression data and improve interpretation of bulk transcriptomics data. However, accuracy of the estimated cell type proportions reported after deconvolution is unknown, which may have a negative impact on the validity of the conclusions drawn. Here, we present UnBlender, a pipeline that enables respiratory scientists to perform cell type deconvolution and routinely evaluate deconvolution accuracy of their approach. UnBlender allows for custom cell type deconvolution tailored to the research question at hand, using consensus cell type labels and validating the approach to promote accurate, reproducible results.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- ScRNAbox: Empowering Single-Cell RNA Sequencing on High Performance Computing Systems 94%
- scConsensus: combining supervised and unsupervised clustering for cell type identification in single-cell RNA sequencing data 94%
- scMuffin: an R package for disentangling solid tumor heterogeneity from single-cell expression data 93%
Similar papers in this journal
- Fast analysis of Spatial Transcriptomics (FaST): an ultra lightweight and fast pipeline for the analysis of high resolution spatial transcriptomics. 93%
- Kmerator Suite: design of specific k-mer signatures andautomatic metadata discovery in large RNA-Seq datasets. 93%
- scROSHI - robust supervised hierarchical identification of single cells 93%
Similar papers in this journal
- Genetic demultiplexing of pooled single-cell RNA-sequencing samples in cancer facilitates effective experimental design 93%
- Extraction of biological terms using large language models enhances the usability of metadata in the BioSample database 93%
- The case for using Mapped Exonic Non-Duplicate (MEND) read counts in RNA-Seq experiments: examples from pediatric cancer datasets 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.