LANTERN: Leveraging Local Ancestry Tracts to Enhance Rare-Variant Aggregate Association Testing
Wang, Y.; Tuftin, B.; Raffield, L. M.; Hidalgo, B.; Kerns, S. L.; DeWan, A. T.; Leal, S. M.; Auer, P.
Show abstract
Individuals with admixed ancestry comprise a significant proportion of populations of the Americas. Statistical methods have been developed to specifically leverage local ancestry inference to enhance the power and interpretability of genome-wide association studies in admixed populations. However, no such methods currently exist to test for rare-variant aggregate associations. Here we present LANTERN (Leveraging local ANcestry Tracts to Enhance Rare variaNt aggregate associations), a method that infers the alleles that lie on each ancestral haplotype and conducts rare-variant aggregate association testing in a generalized linear mixed model framework. Through simulation studies we demonstrated that LANTERN achieves proper control of Type 1 error while boosting power to detect associations when causal alleles predominately lie on one ancestral haplotype. Using data from a cohort of African American participants from the Jackson Heart Study, LANTERN identified two genes known to be involved in red-blood cell (RBC) biology when local ancestry information was incorporated. Specifically, a burden of rare alleles on European ancestral haplotypes in EPO was associated with both hemoglobin levels (HGB) and RBC counts, whereas a burden of rare alleles on African ancestral haplotypes in EPB42 was associated with HGB and RBC. In summary, LANTERN (i) allows for the identification of ancestry-specific rare-variant associations; and (ii) enhances rare-variant association signals compared to an analysis that ignores local ancestry. LANTERN is implemented in R and is freely available on GitHub.
Matching journals
The top 1 journal accounts for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- A new method for multi-ancestry polygenic prediction improves performance across diverse populations 96%
- Scalable generalized linear mixed model for region-based association tests in large biobanks and cohorts 96%
- Leveraging a machine learning derived surrogate phenotype to improve power for genome-wide association studies of partially missing phenotypes in population biobanks 95%
Similar papers in this journal
- A novel method for an unbiased estimate of cross-ancestry genetic correlation using individual-level data 96%
- Population-specific causal disease effect sizes in functionally important regions impacted by selection 96%
- Calibrated rare variant genetic risk scores for complex disease prediction using large exome sequence repositories 95%
Similar papers in this journal
- Exploiting Family History in Aggregation Unit-based Genetic Association Tests 94%
- An expanded analysis framework for multivariate GWAS connects inflammatory biomarkers to functional variants and disease 93%
- An application of the MR-Horse method to reduce selection bias in genome-wide association studies of disease progression 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.