PRS-GRID: A Cross and Within Ancestry Polygenic Risk Prediction Method Based on Individual Genetic Distance
Tang, L.; You, C.; Kong, X.-J.; Napolioni, V.; Jie, H.
Show abstract
BackgroundTwo decades of genome-wide association studies (GWAS) have led to the fast-growing application of polygenic risk prediction (PRS). However, due to population structure and evolutionary path differences, the PRS substrate derived mostly from studies of European ancestry does not work equally well for other ancestries. There is an association between prediction accuracy decay and individual genetic distance (GD) to the genetic centers (GC) of various populations. ObjectivesTo develop a new PRS method and software that utilizes individual GD to improve PRS risk prediction accuracy, especially for non-European populations. MethodWe hypothesize that adding a GD-based weight into PRS methods would enhance its risk prediction performance, particularly for minority groups. We explore the GD first by principal components (PC) and then by phylogenetic tree structures. Building on top of an emerging software (PRS-CSx) that achieves high prediction accuracy across multiple-ancestries, we present PGS-GRID, where "GRID" stands for "Genetic Reference based on Individual Distance". ResultsWe developed a preliminary version of PRS-GRID and pilot tested its prediction performance for a classic quantitative trait (e.g., height) and a disease trait (e.g., type-2 diabetes). We found slight but noticeable improvement of risk prediction in minority populations. We further explored a random forest approach so that the performance of PRS-GRID could be clearly explained, which is a key step for PRS to be used in clinical and public health practice. ConclusionsThe PRS-GRID philosophy and method represent an innovative and significant advancement in the field of polygenic risk prediction. Our work provides a foundation for future research and clinical applications aimed at reducing health disparities and improving population health through personalized medicine.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Exploiting Family History in Aggregation Unit-based Genetic Association Tests 94%
- A Tool for Translating Polygenic Scores onto the Absolute Scale Using Summary Statistics 93%
- Lifestyle Risk Score for aggregating multiple lifestyle factors: Handling missingness of individual lifestyle components in meta-analysis of gene-by-lifestyle interactions 92%
Similar papers in this journal
- A reference panel for linkage disequilibrium and genotype imputation using whole-genome sequencing data from 2,680 participants across India 94%
- Evaluating Genomic Polygenic Risk Scores for Childhood Acute Lymphoblastic Leukemia in Latinos 94%
- Inclusion of Variants Discovered from Diverse Populations Improves Polygenic Risk Score Transferability 94%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.