Molecular Characterization of T-Lineage Acute Lymphoblastic Leukemia by an Optimal-Transport Based Multi-Omics Integration Framework
Li, L.; Wang, J.; Wan, S.
Show abstract
T-lineage acute lymphoblastic leukemia (T-ALL) is an aggressive pediatric malignancy characterized by complex heterogeneity across multiple molecular layers. Accurate subtyping is essential for understanding disease mechanisms, risk stratification, and guiding targeted therapeutic strategies. However, current diagnostic approaches are labor-intensive and time-consuming, and existing computational methods are limited by reliance on single-modality data or simple integration strategies. Effective integration of heterogeneous multi-omics data remains a major computational challenge. We present OTTER (Optimal Transport-based Transcriptomics and gEnomics Representation fusion), a novel multi-modal deep learning framework that jointly models RNA-seq gene expression and somatic genomic variant data for T-ALL molecular characterization. OTTER encodes each omics modality through a modality-specific variational autoencoder and aligns the resulting latent representations using Gromov-Wasserstein optimal transport (GW-OT), which preserves the internal geometric structure of each modality without requiring a shared feature space. We applied OTTER to the Childrens Oncology Group (COG) AALL0434 cohort comprising 1,309 patients across 17 T-ALL subtypes. Gradient-based feature importance and cross-omics interaction analysis were performed on the holdout set to identify subtype-driving molecular features and cross-modal coordinated programs. OTTER provides a principled, biologically interpretable, and computationally effective framework for multi-omics-driven T-ALL molecular characterization. By leveraging GW-OT for geometry-preserving cross-modal alignment and gradient-based interpretability for cross-omics interaction profiling, OTTER goes beyond single-modality approaches to uncover the coordinated molecular landscape of T-ALL. The framework is generalizable to other cancers and multi-omics integration tasks.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- scDREAMER: atlas-level integration of single-cell datasets using deep generative model paired with adversarial classifier 95%
- CellFM: a large-scale foundation model pre-trained on transcriptomics of 100 million human cells 94%
- OmicVerse: A single pipeline for exploring the entire transcriptome universe 94%
Similar papers in this journal
- Inference of cell state transitions and cell fate plasticity from single-cell with MARGARET 94%
- CelLink: integrating single-cell multi-omics data with weak feature linkage and imbalanced cell populations 94%
- CSsingle: A Unified Tool for Robust Decomposition of Bulk and Spatial Transcriptomic Data Across Diverse Single-Cell References 94%
Similar papers in this journal
- An in-depth comparison of linear and non-linear joint embedding methods for bulk and single-cell multi-omics 95%
- SHEST: Single-cell-level artificial intelligence from haematoxylin and eosin morphology for cell type prediction and spatial transcriptomics reconstruction 94%
- PheCode-guided multi-modal topic modeling of electronic health records improves disease incidence prediction and GWAS discovery from UK Biobank 94%
Similar papers in this journal
- Inferring spatial single-cell-level interactions through interpreting cell state and niche correlations learned by self-supervised graph transformer 94%
- Delineating the Effective Use of Self-Supervised Learning in Single-Cell Genomics 93%
- COSIME: Cooperative multi-view integration with Scalable and Interpretable Model Explainer 93%
Similar papers in this journal
- Integrating temporal single-cell gene expression modalities for trajectory inference and disease prediction 94%
- Effective methods for bulk RNA-Seq deconvolution using scnRNA-Seq transcriptomes 94%
- Knowledge-primed neural networks enable biologically interpretable deep learning on single-cell sequencing data 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.