MIDAS: a deep generative model for mosaic integration and knowledge transfer of single-cell multimodal data
He, Z.; Chen, Y.; Hu, S.; An, S.; Shi, J.; Liu, R.; Dong, G.; Shi, J.; Zhao, J.; Wang, J.; Zhu, Y.; Bo, X.; Ying, X.
Show abstract
AO_SCPLOWBSTRACTC_SCPLOWRapidly developing single-cell multi-omics sequencing technologies generate increasingly large bodies of multimodal data. Integrating multimodal data from different sequencing technologies, i.e. mosaic data, permits larger-scale investigation with more modalities and can help to better reveal cellular heterogeneity. However, mosaic integration involves major challenges, particularly regarding modality alignment and batch effect removal. Here we present a deep probabilistic framework for the mosaic integration and knowledge transfer (MIDAS) of single-cell multimodal data. MIDAS simultaneously achieves dimensionality reduction, imputation, and batch correction of mosaic data by employing self-supervised modality alignment and information-theoretic latent disentanglement. We demonstrate its superiority to other methods and reliability by evaluating its performance in full trimodal integration and various mosaic tasks. We also constructed a single-cell trimodal atlas of human peripheral blood mononuclear cells (PBMCs), and tailored transfer learning and reciprocal reference mapping schemes to enable flexible and accurate knowledge transfer from the atlas to new data. Applications in mosaic integration, pseudotime analysis, and cross-tissue knowledge transfer on bone marrow mosaic datasets demonstrate the versatility and superiority of MIDAS.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- uniPort: a unified computational framework for single-cell data integration with optimal transport 98%
- scDREAMER: atlas-level integration of single-cell datasets using deep generative model paired with adversarial classifier 98%
- scMODAL: A general deep learning framework for comprehensive single-cell multi-omics data alignment with feature links 98%
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.