BridgePRS: A powerful trans-ancestry Polygenic Risk Score method
Hoggart, C. J.; Choi, S. W.; Garcia-Gonzalez, J.; Souaiaia, T.; Preuss, M.; O'Reilly, P.
Show abstract
Polygenic Risk Scores (PRS) have huge potential to contribute to biomedical research and to a future of precision medicine, but to date their calculation relies largely on Europeanancestry GWAS data. This global bias makes most PRS substantially less accurate in individuals of non-European ancestry. Here we present BridgePRS, a novel Bayesian PRS method that leverages shared genetic effects across ancestries to increase the accuracy of PRS in non-European populations. The performance of BridgePRS is evaluated in simulated data and real UK Biobank (UKB) data across 19 traits in African, South Asian and East Asian ancestry individuals, using both UKB and Biobank Japan GWAS summary statistics. BridgePRS is compared to the leading alternative, PRS-CSx, and two single-ancestry PRS methods adapted for trans-ancestry prediction. PRS trained in the UK Biobank are then validated out-of-cohort in the independent Mount Sinai (New York) BioMe Biobank. Simulations reveal that BridgePRS performance, relative to PRS-CSx, increases as uncertainty increases: with lower heritability, higher polygenicity, greater between-population genetic diversity, and when causal variants are not present in the data. Our simulation results are consistent with real data analyses in which BridgePRS has better predictive accuracy in African ancestry samples, especially in out-of-cohort prediction (into BioMe), which shows a 60% boost in mean R2 compared to PRS-CSx (P = 2 x 10-6). BridgePRS performs the full PRS analysis pipeline, is computationally efficient, and is a powerful method for deriving PRS in diverse and under-represented ancestry populations.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Joint Modeling of Effect Sizes for Two Correlated Traits: Characterizing Trait Properties to Enhance Polygenic Risk Prediction 97%
- Improving polygenic prediction from summary data by learning patterns of effect sharing across multiple phenotypes. 96%
- Scalable probabilistic PCA for large-scale genetic variation data 96%
Similar papers in this journal
- Using the UK Biobank as a global reference of worldwide populations: application to measuring ancestry diversity from GWAS summary statistics 96%
- The predictive capacity of polygenic risk scores for disease risk is only moderately influenced by imputation panels tailored to the target population 95%
- RapidoPGS: A rapid polygenic score calculator for summary GWAS data without a test dataset 95%
Similar papers in this journal
Similar papers in this journal
- A Comprehensive Evaluation of Methods for Mendelian Randomization Using Realistic Simulations and an Analysis of 38 Biomarkers for Risk of Type-2 Diabetes 95%
- Reweighting the UK Biobank to reflect its underlying sampling population substantially reduces pervasive selection bias due to volunteering 92%
- Bias in two-sample Mendelian randomization when using heritable covariable-adjusted summary associations 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.