Admixture mapping identifies complex trait associations with local ancestry in the All of Us Research Program
Gonzalez Rivera, W. G.; Liu, Y.; Ma, N.; Kim, J.; D'Antonio, M.; Dube, U.; Frazer, K. A.; Gymrek, M.
Show abstract
Genetic studies have largely focused on homogeneous populations, limiting our understanding of the genetic architecture of complex traits in admixed individuals. The advent of diverse biobanks like the All of Us Research Program (AoU) and computationally efficient local ancestry inference (LAI) methods now enable admixture mapping (ADM) at biobank scale. Here, we used two orthogonal LAI methods (GNOMIX and FLARE) to characterize local ancestry in the entire AoU v7.1 cohort (n=230,019). We then used GNOMIX labels to identify associations between African (AFR) and Native American (NAT) local ancestry with 29 quantitative traits. We first analyzed All of Us v7.1 data across African (n=49,797) and Admixed American (n=40,327) cohorts, which identified 97 significant local ancestry associations (65 AFR, 32 NAT). These include strong known signals, such as an association between AFR ancestry at the DARC locus and white blood cell traits and between NAT ancestry at the BUD13/APOE5/ZPR1 locus and triglycerides, as well as additional signals not previously associated with local ancestry. We observed that trait associations with AFR local ancestry are largely consistent across the African and Admixed American cohorts, but that several AFR signals reach genome-wide significance exclusively in Admixed Americans. Grouping associations by trait category revealed distinct ancestral patterns: all endocrine, renal, and 75.0% of liver signals were driven by associations with NAT ancestry, whereas white blood cell (90.9%), red blood cell (65.1%), and lipid (66.7%) signals were largely associated with AFR ancestry, possibly reflecting different population-driven environmental exposures throughout history. Finally, we performed a second round of analysis, comprising the largest ADM study to date, on the entire AoU v7.1 cohort in which we pooled individuals from all ancestries. Despite evidence of confounding due to population structure, summary statistics for pooled results showed strong correlation (r>0.98) with those from single ancestry analysis and detected 2.7x fold more signals, most of which passed at least nominal significance and all of which showed consistent effect directions in the single ancestry cohorts. Overall, these results demonstrate the power of using large, admixed cohorts to gain new insights into the relationship between local ancestry and the genetic architecture of medically relevant complex traits.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Polymorphic short tandem repeats make widespread contributions to blood and serum traits 97%
- Analysis across Taiwan Biobank, Biobank Japan and UK Biobank identifies hundreds of novel loci for 36 quantitative traits 97%
- Meta-analysis fine-mapping is often miscalibrated at single-variant resolution 96%
Similar papers in this journal
Similar papers in this journal
- Enrichment analyses identify shared associations for 25 quantitative traits in over 600,000 individuals from seven diverse ancestries 98%
- Characterizing substructure via mixture modeling in large-scale genetic summary statistics 97%
- Interaction molecular QTL mapping discovers cellular and environmental modifiers of genetic regulatory effects 97%
Similar papers in this journal
- A combined polygenic score of 21,293 rare and 22 common variants significantly improves diabetes diagnosis based on hemoglobin A1C levels 96%
- Genome-wide analyses of 200,453 individuals yields new insights into the causes and consequences of clonal hematopoiesis 96%
- Systematic assessment of regulatory effects of human disease variants in pluripotent cells 96%
Similar papers in this journal
- Evaluating genetic-ancestry inference from single-cell transcriptomic datasets 96%
- Polygenic risk score prediction accuracy convergence 95%
- Multivariate adaptive shrinkage improves cross-population transcriptome prediction for transcriptome-wide association studies in underrepresented populations 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.