FAST: a fast and scalable factor analysis for spatially aware dimension reduction of multi-section spatial transcriptomics data
Liu, W.; Zhang, X.; Chai, X.; Fan, Z.; Lin, H.; Chen, J.; Sun, L.; Yu, T.; Yeong, J.; Liu, J.
Show abstract
Biological techniques for spatially resolved transcriptomics (SRT) have advanced rapidly in both throughput and spatial resolution for a single spatial location. This progress necessitates the development of efficient and scalable spatial dimension reduction methods that can handle large-scale SRT data from multiple sections. Here, we developed FAST as a fast and efficient generalized probabilistic factor analysis for spatially aware dimension reduction, which simultaneously accounts for the count nature of SRT data and extracts a low-dimensional representation of SRT data across multiple sections, while preserving biological effects with consideration of spatial smoothness among nearby locations. Compared with existing methods, FAST uniquely models the count data across multiple sections while using a local spatial dependence with scalable computational complexity. Using both simulated and real datasets, we demonstrated the improved correlation between FAST estimated embeddings and annotated cell/domain types. Furthermore, FAST exhibits remarkable speed, with only FAST being applicable to analyze a mouse embryo Stereo-seq dataset with >2.3 million locations in only 2 hours. More importantly, FAST identified the differential activities of immune-related transcription factors between tumor and non-tumor clusters and also predicted a carcinogenesis factor CCNH as the upstream regulator of differentially expressed genes in a breast cancer Xenium dataset.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- Probabilistic embedding, clustering, and alignment for integrating spatial transcriptomics data with PRECAST 99%
- uniPort: a unified computational framework for single-cell data integration with optimal transport 98%
- Learning interpretable cellular and gene signature embeddings from single-cell transcriptomic data 97%
Similar papers in this journal
- STAN, a computational framework for inferring spatially informed transcription factor activity across cellular contexts 98%
- Probabilistic cell/domain-type assignment of spatial transcriptomics data with SpatialAnno 96%
- CelLink: integrating single-cell multi-omics data with weak feature linkage and imbalanced cell populations 96%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.