Leveraging genetic ancestry continuum information to interpolate PRS for admixed populations
Ruan, Y.; Bhukar, R.; Patel, A.; Koyama, S.; Hull, L.; Truong, B.; Hornsby, W.; Zhang, H.; Chatterjee, N.; Natarajan, P.
Show abstract
The relatively low representation of admixed populations in both discovery and fine-tuning individual-level datasets limits polygenic risk score (PRS) development and equitable clinical translation for admixed populations. Under the assumption that the most informative PRS model for a genetically homogeneous sample varies linearly in an ancestry continuum space, we introduce a Genetic Distance-assisted PRS Combination Pipeline for Diverse Genetic Ancestries (DiscoDivas) to interpolate a harmonized PRS for diverse, especially admixed, genetic ancestries, leveraging multiple PRS models fine-tuned within existing samples, which are mostly of single ancestry, and genetic distance. DiscoDivas treats genetic ancestry as a continuous variable and does not require shifting between different models when calculating PRS for different ancestries. We generated PRS with DiscoDivas and the current conventional method, i.e. fine-tuning multiple GWAS PRS using the matched or similar genetic ancestry samples. DiscoDivas generated a harmonized PRS of the accuracy comparable to or higher than the conventional approach, with the greatest advantage exhibited in admixed individuals.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- FiMAP: A Fast Identity-by-Descent Mapping Test for Biobank-scale Cohorts 97%
- Evaluation of Polygenic Prediction Methodology within a Reference-Standardized Framework 96%
- Joint Modeling of Effect Sizes for Two Correlated Traits: Characterizing Trait Properties to Enhance Polygenic Risk Prediction 96%
Similar papers in this journal
- A novel method for an unbiased estimate of cross-ancestry genetic correlation using individual-level data 97%
- Improved analyses of GWAS summary statistics by reducing data heterogeneity and errors 97%
- Theoretical and empirical quantification of the accuracy of polygenic scores in ancestry divergent populations 96%
Similar papers in this journal
- Identity-by-descent mapping using multi-individual IBD with genome-wide multiple testing adjustment 97%
- GxE PRS: Genotype-environment interaction in polygenic risk score models for quantitative and binary traits 95%
- Meta-MultiSKAT: Multiple phenotype meta-analysis for region-based association test 94%
Similar papers in this journal
- Multivariate adaptive shrinkage improves cross-population transcriptome prediction for transcriptome-wide association studies in underrepresented populations 96%
- Inclusion of Variants Discovered from Diverse Populations Improves Polygenic Risk Score Transferability 96%
- Leveraging Global Genetics Resources to Enhance Polygenic Prediction Across Ancestrally Diverse Populations 96%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.