ECLIPSER: identifying causal cell types and genes for complex traits through single cell enrichment of e/sQTL-mapped genes in GWAS loci
Rouhana, J. M.; Wang, J.; Eraslan, G.; Anand, S.; Hamel, A. R.; Cole, B.; Regev, A.; Aguet, F.; Ardlie, K. G.; Segre, A. V.
Show abstract
SummaryECLIPSER was developed to identify pathogenic cell types and cell type-specific genes that may affect complex disease susceptibility and trait variation by integrating single cell data with known GWAS loci. ECLIPSER maps genes to GWAS loci for a given complex trait based on expression and splicing quantitative trait loci (e/sQTLs) and other functional data, and tests whether the mapped genes are enriched for cell type-specific expression in particular cell types using single-cell/nucleus RNA-seq data from one or more tissues of interest. A Bayesian Fishers exact test is used to compute fold-enrichment significance. We demonstrate the application of ECLIPSER on various skin diseases and traits using snRNA-seq of healthy human skin samples. Availability and ImplementationThe source code and documentation for ECLIPSER and a Jupyter notebook for generating output tables and figures are available at https://github.com/segrelabgenomics/ECLIPSER. The source code for GWASvar2gene that maps genes to GWAS loci based on e/sQTLs is available at https://github.com/segrelabgenomics/GWASvar2gene. The analysis presented here used data from GTEx (https://gtexportal.org/home/datasets) and Open Targets Genetics (https://genetics-docs.opentargets.org/data-access/graphql-api), but can also be applied to other GWAS variant lists and QTL studies. Data used to reproduce the results of the paper are available in Supplementary data.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Set-based rare variant association tests for biobank scale sequencing data sets 93%
- Leveraging functional genomic annotations and genome coverage to improve polygenic prediction of complex traits within and between ancestries 93%
- Adjusting for Common Variant Polygenic Scores Improves Yield in Rare Variant Association Analyses 93%
Similar papers in this journal
- SparkINFERNO: A scalable high-throughput pipeline for inferring molecular mechanisms of non-coding genetic variants 94%
- echolocatoR: an automated end-to-end statistical and functional genomic fine-mapping pipeline 94%
- The predictive capacity of polygenic risk scores for disease risk is only moderately influenced by imputation panels tailored to the target population 94%
Similar papers in this journal
- Integrating Comprehensive Functional Annotations to Boost Power and Accuracy in Gene-Based Association Analysis 95%
- Bayesian multivariate reanalysis of large genetic studies identifies many new associations 94%
- Enhancing Portability of Trans-Ancestral Polygenic Risk Scores through Tissue-Specific Functional Genomic Data Integration 94%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.