Back

Statistical tests for bivariate spatial association across multi-omics data with disjoint coordinates

Hawinkel, S.; Hu, W.; Velten, B.; Maere, S.

2026-06-24 bioinformatics
10.64898/2026.06.19.732250 bioRxiv
Show abstract

Spatial biology has entered a new era of multimodal profiling, with multiple, high-dimensional spatial omics types being measured on consecutive tissue slices, or co-assayed on the same slice. Interest then lies in statistical testing for spatial association between the features of the different modalities, to gain insight in biological processes. One major challenge is the multitude of bivariate combinations, leading to high computational demands. Another difficulty is the difference in spatial resolution between technologies, implying no one-to-one matching between the measurement spots of the two modalities, even after alignment. As a result, common statistical measures such as joint distributions and correlations are not defined, and tests need to rely on spatial vicinity only. Moreover, we argue that many existing bivariate association tests address an inappropriate null hypothesis, or make inappropriate assumptions, both implying absence of spatial autocorrelation in any of the features and leading to misleading conclusions. As a remedy, we modify tests for the detection of spatially variable genes (Morans I, Gaussian processes and generalized additive models (splines)) to derive bivariate tests across modalities with non-overlapping coordinate sets and provide variance estimators that do account for spatial autocorrelation. We develop inference methods for single sections as well as for replicated experiments with multiple sections, and compare their performance in nonparametric and parametric simulations. Finally, we apply the newly developed methods to two co-assayed spatial transcriptomics and metabolomics datasets from mouse and human. The full suite of tests is available from github.com/sthawinke/sbivar as the R-package sbivar. Author summarySpatial biology investigates molecules and cells in their spatial context in living tissues. For this purpose, recent technologies allow to detect different classes of biomolecules on the same or adjacent tissue sections. In this work, we focus on methods to detect colocalization of molecule pairs, which may provide clues on their biological function, e.g. involvement in the same pathways. Many existing statistical tests for identifying such colocalized pairs enforce spot-to-spot matching between tissue sections, thus failing to exploit the full spatial resolution of the data. Moreover, existing methods ignore spatial patterning or autocorrelation naturally present in spatial omics data, leading to many false findings. As a remedy, we present a suite of scalable methods that work directly on disjoint coordinate sets and do account for spatial autocorrelation, confirming their validity in simulation studies. Next we apply the new methods to metabolite and gene expression measurements of mouse and human datasets, uncovering interesting new colocalized gene-metabolite pairs. The methods presented are available in the sbivar R-package from github.com/sthawinke/sbivar.

Matching journals

The top 3 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.