MAGE: Monte Carlo method for Aberrant Gene Expression
Beltran, M.; Joh, R. I.
Show abstract
Identifying genes that are aberrantly expressed is an important first step in the diagnosis and treatment of many diseases. Conventionally, differential expression (DE) analysis is used to screen gene expression profiles to identify functionally associated genes. DE often relies on the variance and fold change in expression from individual genes, which does not consider the expression of all other genes within the profile. When the overall gene expression is skewed, DE does not capture outliers in gene expression. To address this, we have developed a non-parametric DE method based on the probability density for an entire expression profile to select genes that deviate from the global distribution between two gene expression profiles with multiple replicates. Rather than assuming a particular distribution of expression per gene, our method assumes that aberrantly expressed genes (AEGs) will exhibit expression patterns distinguishable from non-AGEs which make up the majority of the profile. Here we introduce our nonparametric method (MAGE: Monte Carlo method for aberrant gene expression) and demonstrate that MAGE can identify AEGs that are not found by conventional DE analyses. The main feature of MAGE is (1) identifying outliers based on the expression profile of all genes rather than performing DE analyses on a per-gene basis and (2) consideration of the variance in expression between two different conditions. We also compared our results with traditional DE analysis as well as density-based clustering methods. MAGE produces consistent results in a variety of conditions and performs conservatively with the addition of noise. We also applied MAGE to single-cell RNA-seq samples and demonstrated that the analysis is robust with subsampling.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- GEOlimma: Differential Expression Analysis and Feature Selection Using Pre-Existing Microarray Data 96%
- A Markov Random Field Model for Network-based Differential Expression Analysis of Single-cell RNA-seq Data 96%
- scTensor detects many-to-many cell-cell interactions from single cell RNA-sequencing data 96%
Similar papers in this journal
- A Generalized Higher-order Correlation Analysis Framework for Multi-Omics Network Inference 96%
- Detection of genes with differential expression dispersion unravels the role of autophagy in cancer progression 96%
- CoVar: A generalizable machine learning approach to identify the coordinated regulators driving variational gene expression 96%
Similar papers in this journal
Similar papers in this journal
- CoRegNet: Unraveling Gene Co-regulation Networks from Public RNA-Seq Repositories Using a Beta-Binomial Statistical Model 96%
- SSMD: A semi-supervised approach for a robust cell type identification and deconvolution of mouse transcriptomics data 96%
- SMNN: Batch Effect Correction for Single-cell RNA-seq data via Supervised Mutual Nearest Neighbor Detection 96%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.