SMaSH: A scalable, general marker gene identification framework for single-cell RNA sequencing and Spatial Transcriptomics
Nelson, M. E.; Riva, S. G.; Cvejic, A.
Show abstract
Spatial transcriptomics is revolutionising the study of single-cell RNA and tissue-wide cell heterogeneity, but few robust methods connecting spatially resolved cells to so-called marker genes from single-cell RNA sequencing, which generate significant insight gleaned from spatial methods, exist. Here we present SMaSH, a general computational framework for extracting key marker genes from single-cell RNA sequencing data for spatial transcriptomics approaches. SMaSH extracts robust and biologically well-motivated marker genes, which characterise the given data-set better than existing and limited computational approaches for global marker gene calculation.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- On the importance of data transformation for data integration in single-cell RNA sequencing analysis 96%
- scEvoNet: a gradient boosting-based method for prediction of cell state evolution 95%
- Benchmarking imputation methods for network inference using a novel method of synthetic scRNA-seq data generation 95%
Similar papers in this journal
- Coffee: Consensus Single Cell-Type Specific Inference For Gene Regulatory Networks 96%
- Species-Agnostic Transfer Learning for Cross-species Transcriptomics Data Integration without Gene Orthology 96%
- scaLR: a low-resource deep neural network-based platform for single cell analysis and biomarker discovery 96%
Similar papers in this journal
- Adjustments to the reference dataset design improves cell type label transfer 96%
- Improved Predictions Of MHC-Peptide Binding Using Protein Language Models 92%
- A Benchmark of state-of-the-art Deconvolution Methods in Spatial Transcriptomics: Insights from Cardiovascular Disease and Chronic Kidney Disease 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.