SubCellSpace: Automated characterization of subcellular mRNA localization patterns in spatial transcriptomics
Wouters, D.; Alvira Larizgoitia, J. I.; Tilkema, N.; Van Minsel, P.; Alar, C.; Seeuws, N.; Koulalis, I. T.; Lee, M.; Vandermeulen, N.; Adivarahan, S.; Da Cruz, S.; Vandereyken, K.; Moor, A.; Thienpont, B.; Voet, T.; Sifrim, A.
Show abstract
The localized translation of transcripts is a universal phenomenon across biological domains. Many examples of subcellular RNA localization and their functional importance have been described. However, these examples remain anecdotal, and a more systematic genome and cell-type-wide analysis is needed. Current spatial transcriptomic techniques can characterize hundreds to thousands of transcript species at subcellular resolutions, enabling the large-scale investigation of subcellular mRNA localization. Here we describe SubCellSpace, a computational framework to learn general representations of mRNA localization patterns. By embedding observed single-cell subcellular localization patterns (SLPs) to an interpretable latent space, SubCellSpace can detect and statistically infer the presence of SLPs, uncover colocalizing gene-pairs and characterize cellular heterogeneity for pattern-presentation. We benchmark SubCellSpace in both synthetic and real data, showing it can correctly detect previously described apical/basal polarized genes in the enterocytes of mouse small-intestine, as well as encode the enterocytes orientation. Additionally, we provide a tailored spatial transcriptomics validation dataset for benchmarking SLP identification based on transcripts previously described to be enriched near subcellular structures in HEK293T cells. We propose a practical and computationally-efficient classification workflow that automatically detects localized transcript species and quantifies their degree of patterning, while controlling false positive rates. Finally, we showcase SubCellSpace in both supervised and unsupervised settings, to either classify pre-determined SLPs or to explore spatial patterning without specifying pattern types a priori. Automated AI models such as SubCellSpace and their integration in spatial transcriptomics analysis workflows will help characterize previously undiscovered subcellular RNA localization phenomena, providing novel insights into post-transcriptional regulation mechanisms.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- scDREAMER: atlas-level integration of single-cell datasets using deep generative model paired with adversarial classifier 96%
- Reference-free cell-type deconvolution of multi-cellular pixel-resolution spatially resolved transcriptomics data 96%
- Characterizing cell-type spatial relationships across length scales in spatially resolved omics data 95%
Similar papers in this journal
- geneBasis: an iterative approach for unsupervised selection of targeted gene panels from scRNA-seq. 96%
- scAlign: a tool for alignment, integration and rare cell identification from scRNA-seq data 96%
- Explainable multi-view framework for dissecting inter-cellular signaling from highly multiplexed spatial data 96%
Similar papers in this journal
- Learning unsupervised feature representations for single cell microscopy images with paired cell inpainting 96%
- Randomized Spatial PCA (RASP): a computationally efficient method for dimensionality reduction of high-resolution spatial transcriptomics data 96%
- Non-linear Archetypal Analysis of Single-cell RNA-seq Data by Deep Autoencoders 95%
Similar papers in this journal
- Cell type identification in spatial transcriptomics data can be improved by leveraging cell-type-informative paired tissue images using a Bayesian probabilistic model. 96%
- noisyR: Enhancing biological signal in sequencing datasets by characterising random technical noise 95%
- CelLink: integrating single-cell multi-omics data with weak feature linkage and imbalanced cell populations 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.