Back

On the use of variational autoencoders for biomedical data integration

Pielies Avelli, M.; Hernandez Medina, R.; Webel, H. E.; Rasmussen, S.

2025-08-22 bioinformatics
10.1101/2025.08.18.670835 bioRxiv
Show abstract

Variational Autoencoders (VAEs) are a widely used framework to integrate diverse biomedical data modalities, create representations that capture the underlying structure of the datasets, and obtain insights about the relations between variables. Here we describe how this is achieved from an empirical point of view in our VAE-based framework MOVE, providing an intuitive perspective on the inner workings of multimodal VAEs in biomedical contexts. We explore how the models emerging dynamics shape their performance and how in silico perturbations can be leveraged to bring to light potential associations between variables. To do that, we extend our framework to handle perturbations of continuous variables, introduce a new approach to better capture associations between them, and create synthetic datasets to benchmark the proposed methods against well defined ground truth associations. We finally showcase our findings in a real biomedical scenario using a multimodal dataset of inflammatory bowel disease.

Matching journals

The top 3 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.