Flexynesis: A deep learning framework for bulk multi-omics data integration for precision oncology and beyond
Uyar, B.; Savchyn, T.; Wurmus, R.; Sarigun, A.; Shaik, M. M.; Franke, V.; Akalin, A.
Show abstract
Accurate decision making in precision oncology depends on integration of multimodal molecular information, such as the genetic data, gene expression, protein abundance, and epigenetic measurements. Deep learning methods facilitate integration of heterogeneous datasets. However, almost all published deep learning-based bulk multi-omics integration methods have constrained usability. They suffer from lack of transparency, modularity, deployability, and are applicable exclusively to narrow tasks. To address these limitations, we introduce Flexynesis, a versatile tool designed with usability, and adaptability in mind. Flexynesis streamlines data processing, enforces structured data splitting, and ensures rigorous model evaluation. It offers unsupervised feature selection, different omics layer fusion options, and hyperparameter tuning. Users can choose from distinct architectures - fully connected networks, variational autoencoders, multi-triplet networks, graph neural networks, and cross-modality encoding networks. Each model is complemented with a straightforward input interface and standardized training, evaluation, and feature importance quantification methods, enabling easy incorporation into data integration pipelines. For improved user experience, Flexynesis supports features such as on-the-fly task determination and compatibility with regression, classification, and survival modeling. It accommodates multi-task prediction of a mixture of numerical/categorical outcome variables with a tolerance for missing labels. We also developed an extensive benchmarking pipeline, showcasing the tools capability across diverse real-life datasets. This toolset should make deep-learning based bulk multi-omics data integration in the context of clinical/pre-clinical data analysis and marker discovery more accessible to a wider audience with or without experience in deep-learning development. Flexynesis is available at https://github.com/BIMSBbioinfo/flexynesis and can be installed from https://pypi.org/project/flexynesis/. O_FIG O_LINKSMALLFIG WIDTH=199 HEIGHT=200 SRC="FIGDIR/small/603606v1_ufig1.gif" ALT="Figure 1"> View larger version (53K): org.highwire.dtl.DTLVardef@17eda85org.highwire.dtl.DTLVardef@13c89ecorg.highwire.dtl.DTLVardef@182f706org.highwire.dtl.DTLVardef@127d9d9_HPS_FORMAT_FIGEXP M_FIG O_FLOATNOGraphical Abstract:C_FLOATNO Summary of the Flexynesis data integration and analysis workflow. C_FIG
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Novel multi-omics deconfounding variational autoencoders can obtain meaningful disease subtyping 97%
- Computationally scalable regression modeling for ultrahigh-dimensional omics data with ParProx 96%
- scaLR: a low-resource deep neural network-based platform for single cell analysis and biomarker discovery 96%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.