S4-Multi: enhancing polygenic score prediction in ancestrally diverse populations
Baierl, J. D.; Tyrer, J. P.; Lai, P.-H.; Gayther, S. A.; Hsiao, Y.-W.; Jones, M. R.; Peng, P.-C. R.; Pharoah, P.
Show abstract
While polygenic scores (PGSs) have shown promise in advancing precision medicine by capturing the additive effects of common germline variants on inherited disease risk, they are presently limited by reduced performance outside of European-origin populations. We extend our previously developed Bayesian polygenic model (PGM) method, select and shrink with summary statistics (S4), to improve prediction accuracy in ancestrally diverse populations. We benchmark this multi-ancestry extension (S4-Multi) against alternative methods on both simulated and biobank data predicting type 2 diabetes, breast cancer, colorectal cancer, asthma, and stroke. In simulation tests, we find that S4-Multi achieves 169% improvement on average over its single ancestry S4 counterpart at prediction in non-European target populations. S4-Multi matches or exceeds top performing methods across the ancestry continuum. In biobank tests, we find that the top-performing PGM method varies considerably by target ancestry and phenotype, with S4-Multi achieving comparable performance to top multi-ancestry methods overall. However, S4-Multi does so while including between 9% and 77% fewer genetic variants relative to competing models, suggesting potential for robust performance in clinical settings with limited available genomic data.
Matching journals
The top 2 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- The Construction of Multi-ethnic Polygenic Risk Score using Transfer Learning 97%
- Evaluating Multi-Ancestry Genome-Wide Association Methods: Statistical Power, Population Structure, and Practical Implications 97%
- CADET: Enhanced transcriptome-wide association analyses in admixed samples using eQTL summary data 97%
Similar papers in this journal
Similar papers in this journal
- MUSSEL: Enhanced Bayesian Polygenic Risk Prediction Leveraging Information across Multiple Ancestry Groups 97%
- Integrative polygenic risk score improves the prediction accuracy of complex traits and diseases 97%
- Incorporating family history of disease improves polygenic risk scores in diverse populations 95%
Similar papers in this journal
- Deep transfer learning provides a Pareto improvement for multi-ancestral clinico-genomic prediction of diseases 95%
- Validation of a Trans-Ancestry Polygenic Risk Score for Type 2 Diabetes in Diverse Populations 94%
- Identifying latent genetic interactions in genome-wide association studies using multiple traits 93%
Similar papers in this journal
- A new method for multi-ancestry polygenic prediction improves performance across diverse populations 98%
- Leveraging functional genomic annotations and genome coverage to improve polygenic prediction of complex traits within and between ancestries 96%
- Computationally efficient whole genome regression for quantitative and binary traits 96%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.