Extending Genome-Wide Association Studies to admixed cohorts with high degrees of relatedness
Tan, T.; Vergara-Lope, A.; Martinez-Magana, J. J.; Shah, N. N.; Yuan, K.; Berumen, J.; Alegre-Diaz, J.; Kuri-Morales, P.; Tapia-Conyer, R.; Gelenter, J.; Montalvo-Ortiz, J. L.; Zhou, W.; Torres, J. M.; Atkinson, E. G.
Show abstract
Recently admixed populations comprise a large portion of the human population worldwide, but are often excluded from Genome-Wide Association Studies (GWAS) due to analytic challenges. Our group has previously developed a local ancestry informed generalized linear model based method, Tractor, for GWAS in admixed samples, which produces accurate ancestry-specific effect sizes and boosts discovery power to identify ancestry-enriched loci. Tractor has been instrumental for elucidating the genetic architecture of complex traits across admixed cohorts, however it operates under an assumption of unrelated samples. As biobanks and other large-scale data sources continue to grow, increasing numbers of closely or cryptically related admixed samples are included. This brings new statistical challenges in conducting GWAS and motivates the timely development of novel tools that can model admixture in various cohort settings. Here, we propose a novel mixed model method, Tractor-Mix, that allows for well-calibrated association studies in datasets containing admixed samples with high degrees of relatedness. Similar to Tractor, our method conducts genetic association tests by leveraging local ancestry to produce more accurate effect sizes and boost power under heterogeneity while effectively controlling false positives. Extensive simulations show this enhanced method is competitive with other state-of-the-art approaches that do not produce ancestry-specific results. Empirical testing of Tractor-Mix on multiple cohorts, including admixed samples from the UK Biobank and Mexico City Prospective Study, highlight the value of the method, identifying ancestry-specific associations. In summary, Tractor-Mix is a powerful association framework that extends the capabilities of current models and will facilitate the inclusion of admixed samples in large-scale GWAS.
Matching journals
The top 2 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Leveraging functional genomic annotations and genome coverage to improve polygenic prediction of complex traits within and between ancestries 98%
- Leveraging fine-mapping and non-European training data to improve trans-ethnic polygenic risk scores 98%
- A resource-efficient tool for mixed model association analysis of large-scale data 98%
Similar papers in this journal
- Characterizing substructure via mixture modeling in large-scale genetic summary statistics 98%
- Enrichment analyses identify shared associations for 25 quantitative traits in over 600,000 individuals from seven diverse ancestries 98%
- Localizing components of shared transethnic genetic architecture of complex traits from GWAS summary data 98%
Similar papers in this journal
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.