SNP heritability: What are we estimating?
Rawlik, K.; Canela-Xandri, O.; Woolliams, J.; Tenesa, A.
Show abstract
The SNP heritability [Formula] has become a central concept in the study of complex traits. Estimation of [Formula] based on genomic variance components in a linear mixed model using restricted maximum likelihood has been widely adopted as the method of choice were individual level data are available. Empirical results have suggested that this approach is not robust if the population of interest departs from the assumed statistical model. Prolonged debate of the appropriate model choice has yielded a number of approaches to account for frequency- and linkage disequilibrium dependent genetic architectures. Here we analytically resolve the question of how these estimates relate to [Formula] of the population from which samples are drawn. In particular, we show that the correct model for the purpose of inference about [Formula] does not require knowledge of the true genetic architecture of a trait. More generally, our results provide a complete perspective of these class of estimators of [Formula], highlighting practical shortcomings of current practise. We illustrate our theoretical results using simulations and data from UK Biobank.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- Beyond SNP Heritability: Polygenicity and Discoverability of Phenotypes Estimated with a Univariate Gaussian Mixture Model 96%
- Estimating indirect parental genetic effects on offspring phenotypes using virtual parental genotypes derived from sibling and half sibling pairs 96%
- Eliciting priors and relaxing the single causal variant assumption in colocalisationanalyses 95%
Similar papers in this journal
Similar papers in this journal
- Probabilistic inference of the genetic architecture underlying functional enrichment of complex traits 97%
- Simultaneous estimation of bi-directional causal effects and heritable confounding from GWAS summary statistics 97%
- Fast Kernel-based Association Testing of non-linear genetic effects for Biobank-scale data 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.