Statistical uncertainty explains the poor agreement in polygenic scoring for type 2 diabetes
Mandla, R.; Li, X.; Shi, Z.; Abramowitz, S.; Lapinska, S.; Penn Medicine Biobank, ; Levin, M. G.; Damrauer, S. M.; Pasaniuc, B.
Show abstract
Polygenic scores (PGS) have emerged as an important tool for genetic risk prediction in medicine to identify individuals at high-risk for disease. A major limitation in their implementation is the apparent disagreement among scores for the same individual decreasing their interpretability and utility in clinical settings. Here we show that the poor agreement across PGSes for type 2 diabetes (T2D) is fully explained by statistical uncertainty in PGS-based prediction; individual-level uncertainty estimates from a single PGS explain the variability across existing PGSes. We provide an approach for the selection of high-risk individuals that incorporates measures of uncertainty and show that individuals with high confidence based on their PGS uncertainty have higher risk agreement across existing PGS and are more likely to develop T2D than high-risk individuals based on only point estimates of PGS. Together, these findings shed light on the factors underlying a roadblock in PGS implementation and underscore the need to incorporate uncertainty in PGS-based predictions.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Transferability of genetic loci and polygenic scores for cardiometabolic traits in British Pakistanis and Bangladeshis 97%
- Calibrated rare variant genetic risk scores for complex disease prediction using large exome sequence repositories 97%
- Common genetic variation associated with Mendelian disease severity revealed through cryptic phenotype analysis 96%
Similar papers in this journal
- A combined polygenic score of 21,293 rare and 22 common variants significantly improves diabetes diagnosis based on hemoglobin A1C levels 97%
- Combining case-control status and family history of disease increases association power 97%
- Inferring compound heterozygosity from large-scale exome sequencing data 97%
Similar papers in this journal
- Analysis across Taiwan Biobank, Biobank Japan and UK Biobank identifies hundreds of novel loci for 36 quantitative traits 97%
- Integrative polygenic risk score improves the prediction accuracy of complex traits and diseases 97%
- Incorporating family history of disease improves polygenic risk scores in diverse populations 97%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.