eMAGMA: An eQTL-informed method to identify risk genes using genome-wide association study summary statistics
Gerring, Z. F.; Mina-Vargas, A.; Derks, E. M.
Show abstract
Identifying genes underlying genetic associations of complex disease is challenging because most common risk variants reside in non-protein coding regions of the genome and likely alter the expression of target genes by disrupting tissue and cell-type specific regulatory elements. To address this challenge, we developed a methodological framework, eQTL-MAGMA (eMAGMA), that converts SNP-level summary statistics into gene-level association statistics by assigning non-coding SNPs to their putative genes based on tissue-specific eQTL information. We compared eMAGMA to three eQTL informed gene-based approaches--S-PrediXcan, FUSION, and SMR--using simulated phenotype data. Phenotypes were simulated based on eQTL reference data using GCTA for all genes with at least one eQTL at chromosome 1 (651 genes). We performed 10 simulations per gene. The eQTL-h2 (i.e., the proportion of variation explained by the eQTLs was set at 1%, 2%, and 5%. We found eMAGMA outperforms other gene-based approaches across a range of simulated parameters (e.g. the number of identified causal genes). When applied to genome-wide association summary statistics for major depression, eMAGMA identified substantially more putative candidate causal genes compared to other eQTL-based approaches. By integrating tissue-specific eQTL information, these results show eMAGMA will help to identify novel candidate causal genes from genome-wide association summary statistics and thereby improve the understanding of the biological basis of complex disorders.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Integration of multidimensional splicing data and GWAS summary statistics for risk gene discovery 95%
- Joint Modeling of Effect Sizes for Two Correlated Traits: Characterizing Trait Properties to Enhance Polygenic Risk Prediction 95%
- Leveraging expression from multiple tissues using sparse canonical correlation analysis (sCCA) and aggregate tests improves the power of transcriptome-wide association studies (TWAS) 95%
Similar papers in this journal
- BinomiRare: A carriers-only test for association of rare genetic variants with a binary outcome for mixed models and any case-control proportion 94%
- Pitfalls in performing genome-wide association studies on ratio traits 93%
- Inverted genomic regions between reference genome builds in humans impact imputation accuracy and decrease the power of association testing 93%
Similar papers in this journal
- TWAS pathway method greatly enhances the number of leads for uncovering the molecular underpinnings of psychiatric disorders 94%
- Schizophrenia Risk Alleles Often Affect The Expression of Many Genes and Each Gene May Have a Different Effect On The Risk; A Mediation Analysis. 93%
- Improving Machine Learning Prediction of ADHD Using Gene Set Polygenic Risk Scores and Risk Scores from Genetically Correlated Phenotypes 91%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.