Move BeTween modAlities (MBTA) employs flow matching to predict single cell data modalities
Xu, B.; Zhang, Y.; Michor, F.
Show abstract
Integrating diverse molecular modalities to obtain a comprehensive view of cellular identity remains a major challenge in single-cell biology. A fundamental but underappreciated obstacle is structural mismatch -- the phenomenon in which the neighborhood structure of a cell differs depending on which molecular modality is used to define it. Existing approaches typically embed modalities into a shared latent space, which actively erases the structural differences between modalities that make multimodal measurements scientifically valuable. Here we introduce Move BeTween modAlities (MBTA), the first framework explicitly designed to address structural mismatch. Rather than forcing modalities into a shared representation, MBTA maintains modality-specific latent spaces and connects them via flow matching, preserving the structural integrity of each modality while enabling accurate cross-modal translation. Across extensive benchmarks on multi-modal single-cell datasets, MBTA consistently outperformed existing methods, with the largest gains observed in datasets with pronounced structural mismatch. Applied to joint genomic and transcriptomic profiles of breast cancer patients, MBTA identified transcriptomic lineage relationships corroborated by genomic variation and outperformed state-of-the-art transcriptomics-based copy number inference methods. Extending this framework to mouse embryonic development, we reconstructed temporal trajectories jointly defined by gene expression and seven complementary epigenetic modalities. MBTA can connect any number of molecular readouts without erasing their individual character, serving as the computational foundation for assembling multi-layered portraits of cells.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- CAPTAIN: A multimodal foundation model pretrained on co-assayed single-cell RNA and protein 98%
- Multi-modal Diffusion Model with Dual-Cross-Attention for Multi-Omics Data Generation and Translation 97%
- DGAT: A Dual-Graph Attention Network for Inferring Spatial Protein Landscapes from Transcriptomics 97%
Similar papers in this journal
- Learning multi-cellular representations of single-cell transcriptomics data enables characterization of patient-level disease states 97%
- TarDis: Achieving Robust and Structured Disentanglement of Multiple Covariates 96%
- Identifying maximally informative signal-aware representations of single-cell data using the Information Bottleneck 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.