Leveraging haplotype information in heritability estimation and polygenic prediction
Meisner, J.; Benros, M. E.; Rasmussen, S.
Show abstract
Polygenic prediction has yet to make a major clinical breakthrough in precision medicine and psychiatry, where the application of polygenic risk scores are expected to improve clinical decision-making. Most widely used approaches for estimating polygenic risk scores are based on summary statistics from external large-scale genome-wide association studies, which relies on assumptions of matching data distributions. This may hinder the impact of polygenic risk scores in modern diverse populations due to small differences in genetic architectures. Reference-free estimators of polygenic scores are instead based on genomic best linear unbiased predictions and models the population of interest directly. We introduce a framework, named hapla, with a novel algorithm for clustering haplotypes in phased genotype data to estimate heritability and perform reference-free polygenic prediction in complex traits. We utilize inferred haplotype clusters to compute accurate SNP heritability estimates and polygenic scores in a simulation study and the iPSYCH2012 case-cohort for depression disorders and schizophrenia. We demonstrate that our haplotype-based approach robustly outperforms standard genotype-based approaches, which can help pave the way for polygenic risk scores in the future of precision medicine and psychiatry. hapla is freely available at https://github.com/Rosemeis/hapla.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Joint Modeling of Effect Sizes for Two Correlated Traits: Characterizing Trait Properties to Enhance Polygenic Risk Prediction 97%
- Beyond SNP Heritability: Polygenicity and Discoverability of Phenotypes Estimated with a Univariate Gaussian Mixture Model 96%
- Improving polygenic prediction from summary data by learning patterns of effect sharing across multiple phenotypes. 95%
Similar papers in this journal
- A new method for multi-ancestry polygenic prediction improves performance across diverse populations 95%
- LDAK-KVIK performs fast and powerful mixed-model association analysis of quantitative and binary phenotypes 95%
- Computationally efficient whole genome regression for quantitative and binary traits 95%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.