Back

Mosaic integration of spatial multi-omics with SpaMosaic

Yan, X.; Li, M.; Ang, K. S.; Olst, L. v.; Edwards, A.; Watson, T.; Zheng, R.; Fan, R.; Gate, D.; Chen, J.

2024-10-03 bioinformatics
10.1101/2024.10.02.616189 bioRxiv
Show abstract

With the advent of spatial multi-omics, we can mosaic integrate diverse datasets with partially overlapping modalities to construct consensus multi-modal spatial atlases of the source tissue. SpaMosaic is a spatial multi-omics mosaic integration tool that employs contrastive learning and graph neural networks to construct a modality-agnostic and batch-corrected latent space suited for analyses like spatial domain identification and imputing missing omes. Using simulated and experimentally acquired datasets, we benchmarked SpaMosaic against single-cell multi-omics mosaic integration methods. The experimental data encompassed RNA and protein abundance, chromatin accessibility or histone modifications, acquired from brain, embryo, tonsil, and lymph node tissues. SpaMosaic achieved superior performance over existing methods in identifying known spatial domains with enhanced resolution and clarity while reducing noise and batch effects. It also ranked top for modality alignment, enabling seamless diagonal integration without the need for image registration. After integration, SpaMosaic can also impute missing modalities. With a mosaic set of mouse brain data with RNA and different epigenomic modalities, we integrated and imputed the missing omics. There we found the imputed gene activity scores of activating and silencing histone marks show the correct correlation with RNA. Moreover, we uncovered more region-specific genes and pathways showing the correct transcriptome-epigenome correlations in the imputed histone modification than in the measured chromatin accessibility modalities. Lastly, SpaMosaics imputation also allows the inference of relationships between different modalities without requiring co-profiling from the same section.

Published in Nature Genetics (predicted rank #22) · training set

Matching journals

The top 3 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.