Back

Statistical uncertainty explains the poor agreement in polygenic scoring for type 2 diabetes

Mandla, R.; Li, X.; Shi, Z.; Abramowitz, S.; Lapinska, S.; Penn Medicine Biobank, ; Levin, M. G.; Damrauer, S. M.; Pasaniuc, B.

2026-02-27 genetic and genomic medicine
10.64898/2026.02.25.26347015 medRxiv
Show abstract

Polygenic scores (PGS) have emerged as an important tool for genetic risk prediction in medicine to identify individuals at high-risk for disease. A major limitation in their implementation is the apparent disagreement among scores for the same individual decreasing their interpretability and utility in clinical settings. Here we show that the poor agreement across PGSes for type 2 diabetes (T2D) is fully explained by statistical uncertainty in PGS-based prediction; individual-level uncertainty estimates from a single PGS explain the variability across existing PGSes. We provide an approach for the selection of high-risk individuals that incorporates measures of uncertainty and show that individuals with high confidence based on their PGS uncertainty have higher risk agreement across existing PGS and are more likely to develop T2D than high-risk individuals based on only point estimates of PGS. Together, these findings shed light on the factors underlying a roadblock in PGS implementation and underscore the need to incorporate uncertainty in PGS-based predictions.

Matching journals

The top 4 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.