miDGD: a multi-modal deep generative model predicts miRNA expression from bulk or single-cell mRNA expression
Zamani, F.; Rasmussen, A. M.; Schuster, V.; Diekema, M. H.; Krogh, A.; Pedersen, J. S.
Show abstract
MicroRNAs (miRNAs) are important post-transcriptional regulators, yet their expression is typically unobserved in single-cell and most bulk RNA-seq datasets. We present miDGD, a deep generative decoder model that predicts miRNA abundance directly from gene expression alone. Trained on bulk and single-cell datasets from TCGA, GTEx, and human cell lines, miDGD learned a shared latent representation of matched mRNA and miRNA profiles that organized samples into biologically meaningful clusters reflecting tissue and cancer types. The model reconstructed both tissue-specific and broadly expressed miRNAs, recapitulated known miRNA-target relationships, and showed robust performance in sparse and single-cell data. miDGD outperformed miRSCAPE and recent miRNA activity inference methods, with improved cross-dataset generalization. These results establish a deep generative model as an improved framework for predicting miRNA expression when direct measurements are unavailable.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- CREaTor: zero-shot cis-regulatory pattern modeling with attention mechanisms 97%
- Characterizing Spatially Continuous Variations in Tissue Microenvironment through Niche Trajectory Analysis 97%
- STHD: probabilistic cell typing of single Spots in whole Transcriptome spatial data with High Definition 96%
Similar papers in this journal
- Coralysis enables sensitive identification of imbalanced cell types and states in single-cell data via multi-level integration 96%
- SPOTlight:Seeded NMF regression to Deconvolute Spatial Transcriptomics Spots with Single-Cell Transcriptomes 96%
- Unraveling the start element and regulatory divergence of core promoters across the domain Bacteria 96%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.