Isoform-level discovery, quantification and fusion analysis from single-cell and spatial long-read RNA-seq data with Bambu-Clump
Sim, A. D.; Ling, M. H.; Chen, Y.; Lu, H.; See, Y. X.; Perrin, A.; Leng Agnes, O. B.; Cao, E. Y.; Chia, B.; Liu, J.; Wüstefeld, T.; Shin, J.; Göke, J.
Show abstract
Single cell and spatial transcriptomics have dramatically changed how we can profile RNA from heterogenous biological samples. Combining single cell and spatial profiling with long read RNA-Seq promises to enable the discovery and quantification of individual RNA isoforms at the single-cell level. However, highly multiplexed data such as from a single cell experiment only generates a limited number of reads for each cell, constituting a major challenge for transcript discovery and quantification with existing approaches that usually have limited power for samples with low sequencing depth. Here we present Bambu-Clump, a computational method that performs transcript discovery and quantification from single cell and spatial long read RNA-Seq data using information from both each cell and the cell cluster. Using this approach, Bambu-Clump provides the most accurate transcript discovery compared to other existing methods, and improves transcript quantification compared to methods that rely on estimates derived from single cells. We apply Bambu-Clump to identify fusion transcripts in single-cells, compare 5 and 3 selection protocols, and identify novel isoform cell-type markers in spatial mouse brain data. Together, Bambu-Clump provides an easy-to-use, efficient, and accurate method for analysing individual isoform expression for single cells and cell clusters across multiple datasets and replicates from long read RNA-Seq.
Matching journals
The top 2 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- High throughput, error corrected Nanopore single cell transcriptome sequencing 97%
- RoCK and ROI: Single-cell transcriptomics with multiplexed enrichment of selected transcripts and region-specific sequencing 97%
- Scarf: A toolkit for memory efficient analysis of large-scale single-cell genomics data 97%
Similar papers in this journal
- Automated quality control and cell identification of droplet-based single-cell data using dropkick 96%
- Highly accurate reference and method selection for universal cross-dataset cell type annotation with CAMUS 96%
- scTIE: data integration and inference of gene regulation using single-cell temporal multimodal data 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.