Genetic Fine-mapping with Dense Linkage Disequilibrium Blocks: genetics of nicotine dependence
Mo, C.; Ye, Z.; Hatch, K.; Zhang, Y.; Wu, Q.; Liu, S.; Kochunov, P.; Hong, L. E.; MA, T.; Chen, S.
Show abstract
Fine-mapping is an analytical step to perform causal prioritization of the polymorphic variants on a trait-associated genomic region observed from genome-wide association studies (GWAS). The prioritization of causal variants can be challenging due to the linkage disequilibrium (LD) patterns among hundreds to thousands of polymorphisms associated with a trait. We propose a novel{ell} 0 graph norm shrinkage algorithm to select causal variants from dense LD blocks consisting of highly correlated SNPs that may not be proximal or contiguous. We extract dense LD blocks and perform regression shrinkage to calculate a prioritization score to select a parsimonious set of causal variants. Our approach is computationally efficient and allows performing fine-mapping on thousands of polymorphisms. We demonstrate its application using a large UK Biobank (UKBB) sample related to nicotine addiction. Our results suggest that polymorphic variances in both neighboring and distant variants can be consolidated into dense blocks of highly correlated loci. Simulations were used to evaluate and compare the performance of our method and existing fine-mapping algorithms. The results demonstrated that our method outperformed comparable fine-mapping methods with increased sensitivity and reduced false-positive error rate regarding causal variant selection. The application of this method to smoking severity trait in UKBB sample replicated previously reported loci and suggested the causal prioritization of genetic effects on nicotine dependency. Author summaryDisentangling the complex linkage disequilibrium (LD) pattern and selecting the underlying causal variants have been a long-term challenge for genetic fine-mapping. We find that the LD pattern within GWAS loci is intrinsically organized in delicate graph topological structures, which can be effectively learned by our novel{ell} 0 graph norm shrinkage algorithm. The extracted LD graph structure is critical for causal variant selection. Moreover, our method is less constrained by the width of GWAS loci and thus can fine-map a massive number of correlated SNPs.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- CoMM-S2: a collaborative mixed model using summary statistics in transcriptome-wide association studies 97%
- LPM: a latent probit model to characterize the relationship among complex traits using summary statistics from multiple GWASs and functional annotations 96%
- Unpaired Data Empowers Association Tests 96%
Similar papers in this journal
- Inferring Causal Direction Between Two Traits in the Presence of Horizontal Pleiotropy with GWAS Summary Data 96%
- A novel method for multiple phenotype association studies based on genotype and phenotype network 96%
- Robust Inference of Bi-Directional Causal Relationships in Presence of Correlated Pleiotropy with GWAS Summary Data 96%
Similar papers in this journal
- netMUG: a novel network-guided multi-view clustering workflow for dissecting genetic and facial heterogeneity 95%
- Boosting polygenic risk scores 94%
- New analysis framework incorporating mixed mutual information and scalable Bayesian networks for multimodal high dimensional genomic and epigenomic cancer data 93%
Similar papers in this journal
- OSCAA: A Two-Dimensional Gaussian Mixture Model for Copy Number Variation Association Analysis 95%
- A Two-Sample Robust Bayesian Mendelian Randomization Method Accounting for Linkage Disequilibrium and Idiosyncratic Pleiotropy with Applications to the COVID-19 Outcome 95%
- DYNATE: Localizing Rare-Variant Association Regions via Multiple Testing Embedded in an Aggregation Tree 94%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.