Signal Strength Aware Latent Spaces Reveal Molecularly Distinct Substructures within Human Kidney Tissue
Delacour, P.-L.; Migas, L.; Farrow, M.; Yang, H.; Fogo, A. B.; Spraggins, J. M.; Van de Plas, R.
Show abstract
As datasets grow increasingly high-dimensional and complex, distinguishing a condensed set of interpretable underlying factors becomes essential. In spatial omics, for example, hundreds to thousands of molecular features per observation promise unprecedented biological insight. However, without meaningful latent representations, that potential remains markedly untapped. We propose a new approach based on the beta-variational autoencoder and kernel density estimation to dissect data along independent, uncertainty-aware, and interpretable (yet non-linear) latent axes. We include a novel comparative-latent-traversal algorithm to translate latent findings back into the original measurement context. Demonstrating on imaging mass spectrometry-based molecular imaging of human kidney, the approachs disentangling properties are shown to impress a latent space structure that separates signal strength from relative signal content, offering exceptional chemical insight. Our approach uncovers unexpected subdivisions within kidney proximal tubules, confirmed to be biological, and reveals hereto-unknown lipid species differentiating them. This confirms our workflows potential as an interpretation-and-hypothesis-generating discovery tool.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Single-Cell Multi-Modal GAN (scMMGAN) reveals spatial patterns in single-cell data from triple negative breast cancer 94%
- MUSTANG: MUlti-sample Spatial Transcriptomics data ANalysis with cross-sample transcriptional similarity Guidance 92%
- Application of Aligned-UMAP to longitudinal biomedical studies 92%
Similar papers in this journal
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.