Large uncertainty in individual PRS estimation impacts PRS-based risk stratification
Ding, Y.; Hou, K.; Burch, K. S.; Lapinska, S.; Prive, F.; Vilhjalmsson, B. J.; Sankararaman, S.; Pasaniuc, B.
Show abstract
Large-scale genome-wide association studies have enabled polygenic risk scores (PRS), which estimate the genetic value of an individual for a given trait. Since PRS accuracy is typically assessed using cohort-level metrics (e.g., R2), uncertainty in PRS estimates at individual level remains underexplored. Here we show that Bayesian PRS methods can estimate the variance of an individuals PRS and can yield well-calibrated credible intervals for the genetic value of a single individual. For real traits in the UK Biobank (N=291,273 unrelated "white British") we observe large variance in individual PRS estimates which impacts interpretation of PRS-based stratification; for example, averaging across 13 traits, only 0.8% (s.d. 1.6%) of individuals with PRS point estimates in the top decile have their entire 95% credible intervals fully contained in the top decile. We provide an analytical estimator for individual PRS variance--a function of SNP-heritability, number of causal SNPs, and sample size--and observe high concordance with individual variances estimated via posterior sampling. Finally as an example of the utility of individual PRS uncertainties, we explore a probabilistic approach to PRS-based stratification that estimates the probability of an individuals genetic value to be above a prespecified threshold. Our results showcase the importance of incorporating uncertainty in individual PRS estimates into subsequent analyses.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- Eliciting priors and relaxing the single causal variant assumption in colocalisationanalyses 96%
- FiMAP: A Fast Identity-by-Descent Mapping Test for Biobank-scale Cohorts 96%
- Leveraging expression from multiple tissues using sparse canonical correlation analysis (sCCA) and aggregate tests improves the power of transcriptome-wide association studies (TWAS) 96%
Similar papers in this journal
Similar papers in this journal
- Hidden structure in polygenic scores and the challenge of disentangling ancestry interactions in admixed populations 96%
- Testing for differences in polygenic scores in the presence of confounding 96%
- Characterization of direct and/or indirect genetic associations for multiple traits in longitudinal studies of disease progression 95%
Similar papers in this journal
- Efficient and Flexible Integration of Variant Characteristics in Rare Variant Association Studies Using Integrated Nested Laplace Approximation 95%
- Improving the coverage of credible sets in Bayesian genetic fine-mapping 95%
- Cross-fitted instrument: a blueprint for one-sample Mendelian Randomization 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.