A Stacking Framework for Polygenic Risk Prediction in Admixed Individuals
Liao, K.; Zöllner, S.
Show abstract
1.1Polygenic risk scores (PRS) are summaries of an individuals personalized genetic risk for a trait or disease. However, PRS often perform poorly for phenotype prediction when the ancestry of the target population does not match the population in which GWAS effect sizes were estimated. For many populations this can be addressed by performing GWAS in the target population. However, admixed individuals (whose genomes can be traced to multiple ancestral populations) lie on an ancestry continuum and are not easily represented as a discrete population. Here, we propose slaPRS (stacking local ancestry PRS), which incorporates multiple ancestry GWAS to alleviate the ancestry dependence of PRS in admixed samples. slaPRS uses ensemble learning (stacking) to combine local population specific PRS in regions across the genome. We compare slaPRS to single population PRS and a method that combines single population PRS globally. In simulations, slaPRS outperformed existing approaches and reduced the ancestry dependence of PRS in African Americans. In lipid traits from African British individuals (UK Biobank), slaPRS again improved on single population PRS while performing comparably to the globally combined PRS. slaPRS provides a data-driven and flexible framework to incorporate multiple population-specific GWAS and local ancestry in samples of admixed ancestry.
Matching journals
The top 2 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Evaluating Multi-Ancestry Genome-Wide Association Methods: Statistical Power, Population Structure, and Practical Implications 98%
- Non-parametric polygenic risk prediction using partitioned GWAS summary statistics 98%
- CADET: Enhanced transcriptome-wide association analyses in admixed samples using eQTL summary data 98%
Similar papers in this journal
- Improved analyses of GWAS summary statistics by reducing data heterogeneity and errors 97%
- Quantifying portable genetic effects and improving cross-ancestry genetic prediction with GWAS summary statistics 97%
- Flashfm: A Flexible and Shared Information Fine-mapping Approach for Multiple Quantitative Traits 97%
Similar papers in this journal
- Inclusion of Variants Discovered from Diverse Populations Improves Polygenic Risk Score Transferability 97%
- Multivariate adaptive shrinkage improves cross-population transcriptome prediction for transcriptome-wide association studies in underrepresented populations 96%
- Evaluating genetic-ancestry inference from single-cell transcriptomic datasets 96%
Similar papers in this journal
- A new method for multi-ancestry polygenic prediction improves performance across diverse populations 98%
- Leveraging functional genomic annotations and genome coverage to improve polygenic prediction of complex traits within and between ancestries 97%
- Leveraging a machine learning derived surrogate phenotype to improve power for genome-wide association studies of partially missing phenotypes in population biobanks 97%
Similar papers in this journal
- FiMAP: A Fast Identity-by-Descent Mapping Test for Biobank-scale Cohorts 97%
- Accurate detection of shared genetic architecture from GWAS summary statistics in the small-sample context 96%
- Adjusting for principal components can induce spurious associations in genome-wide association studies in admixed populations 96%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.