scGAIN: Single Cell RNA-seq Data Imputation using Generative Adversarial Networks
Gunady, M. K.; Kancherla, J.; Bravo, H. C.; Feizi, S.
Show abstract
Single cell RNA sequencing (scRNA-seq) provides a rich view into the heterogeneity underlying a cell population. However single-cell data are usually noisy and very sparse due to the presence of dropout genes. In this work we propose an approach to impute missing gene expressions in single cell data using generative adversarial networks (GANs). By learning an approximate distribution of the data, our approach, scGAIN, can impute dropouts in simulated and real single cell data. The work in this paper discusses how to adopt GAIN training model into the domain of imputing single cell data. Experiments show that scGAIN gives competitive results compared to the state-of-the-art approaches while showing superiority in various aspects in simulation and real data. Imputation by scGAIN successfully recovers the underlying clustering of different subpopulations, provides sharp estimates around true mean expressions and increase the correspondence with matched bulk RNAseq experiments.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- ACTIVA: realistic single-cell RNA-seq generation with automatic cell-type identification using introspective variational autoencoders 97%
- Consensus Label Propagation with Graph Convolutional Networks for Single-Cell RNA Sequencing Cell Type Annotation 97%
- CellVGAE: An unsupervised scRNA-seq analysis workflow with graph attention networks 97%
Similar papers in this journal
Similar papers in this journal
- Graph Contrastive Learning as a Versatile Foundation for Advanced scRNA-seq Data Analysis 97%
- Synthetic observations from deep generative models and binary omics data with limited sample size 97%
- Species-Agnostic Transfer Learning for Cross-species Transcriptomics Data Integration without Gene Orthology 96%
Similar papers in this journal
- Benchmarking imputation methods for network inference using a novel method of synthetic scRNA-seq data generation 95%
- eSVD-DE: Cohort-wide differential expression in single-cell RNA-seq data using exponential-family embeddings 95%
- PyClone-VI: Scalable inference of clonal population structures using whole genome data 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.