Trans-ancestry polygenic models for the prediction of LDL blood levels: An analysis of the UK Biobank and Taiwan Biobank
Hassanin, E.; Lee, K.-H.; Hsieh, T.-C.; Aldisi, R.; Lee, Y.-L.; Bobbili, D.; Krawitz, P.; May, P.; Chen, C.-Y.; Maj, C.
Show abstract
BackgroundPolygenic risk scores (PRSs) are proposed for use in clinical and research settings for risk stratification. PRS predictions often show bias toward the population of available genome-wide association studies, which is typically of European ancestry. This study aims to assess the performance differences of ancestry-specific PRS and test the implementation of multi-ancestry PRS to enhance the generalizability of low-density lipoprotein (LDL) cholesterol predictions in the East Asian population MethodsWe computed ancestry-specific and multi-ancestry PRS for LDL using data from the global lipid consortium while accounting for population-specific linkage disequilibrium patterns using PRS-CSx method. We first conducted an ancestry-wide analysis using the UK Biobank dataset (n=423,596) and then applied the same models to the Taiwan Biobank dataset (TWB, n=68,978). PRS performances were based on linear regression with adjustment for age, sex, and principal components. PRS strata were considered to assess the extent to which a PRS categorization can stratify individuals for LDL cholesterol levels in East Asian samples. ResultsPopulation-specific PRS better predicted LDL levels within the target population but multi-ancestry PRS were more generalizable. In the TWB dataset, covariate-adjusted R2 values were 9.3% for ancestry-specific PRS, 6.7% for multi-ancestry PRS, and 4.5% for European-specific PRS. Similar trends (8.6%, 7.8%, 6.2%) were observed in the smaller East Asian population of the UK Biobank (n=1,480). Consistent with the R2 values, PRS stratification in East Asians (TWB) effectively captured a heterogenous variability in LDL blood cholesterol levels across PRS strata. The mean difference in LDL levels between the lowest and highest East Asian-specific PRS (EAS_PRS) deciles was 0.82, compared to 0.59 for European-specific PRS (EUR_PRS) and 0.76 for multi-ancestry PRS. Notably, the mean LDL values in the top decile of multi-ancestry PRS were comparable to those of EAS_PRS (3.543 vs. 3.541, P=0.86). ConclusionsOur analysis of the PRS prediction model for LDL cholesterol further supports the issue of PRS generalizability across populations. Our targeted analysis of the East Asian (EAS) population revealed that integrating non-European genotyping data, accounting for population-specific linkage disequilibrium, and considering meta-analyses of non-European-based GWAS alongside powerful European-based GWAS can enhance the generalizability of LDL PRS.
Matching journals
The top 9 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Inclusion of Variants Discovered from Diverse Populations Improves Polygenic Risk Score Transferability 93%
- A reference panel for linkage disequilibrium and genotype imputation using whole-genome sequencing data from 2,680 participants across India 93%
- An LDLR missense variant poses high risk of familial hypercholesterolemia in 30% of Greenlanders and offers potential for early cardiovascular disease intervention 93%
Similar papers in this journal
- R2ROC: An efficient method of comparing two or more correlated AUC from out-of-sample prediction using polygenic scores 93%
- Common, low-frequency, rare, and ultra-rare coding variants contribute to COVID-19 severity 90%
- Ancestry-specific polygenic scores and SNP heritability of 25(OH)D in African- and European-ancestry populations 90%
Similar papers in this journal
- Genetic Risk Scores for Cardiometabolic Traits in Sub-Saharan African Populations 95%
- High-throughput multivariable Mendelian randomization analysis prioritizes apolipoprotein B as key lipid risk factor for coronary artery disease 92%
- Educational attainment as a modifier of the effect of polygenic scores for cardiovascular risk factors: cross-sectional and prospective analysis of UK Biobank 92%
Similar papers in this journal
- Exploiting Family History in Aggregation Unit-based Genetic Association Tests 93%
- Lifestyle Risk Score for aggregating multiple lifestyle factors: Handling missingness of individual lifestyle components in meta-analysis of gene-by-lifestyle interactions 93%
- Assessment of ability of AlphaMissense to identify variants affecting susceptibility to common disease 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.