Personalized Feature Statistics: Individual-Level Variant Inference under Genetic Ancestry Continuum
Wang, J. F.; Yu, R.; Edelson, J.; Park, J.; Le Guen, Y.; Liu, X.; Belloy, M.; Ionita-Laza, I.; Greicius, M.; Tang, H.; He, Z.
Show abstract
Genome-wide association studies (GWAS) have successfully identified numerous genetic variants associated with complex diseases. However, the extent to which the effects of these variants vary across populations of diverse ancestries remains poorly understood. Furthermore, in these contexts genetic ancestry is treated as a categorical variable, thereby oversimplifying its continuous nature and the more nuanced ways in which it can influence genetic effects on disease. Here, we propose personalized feature statistics (PFstatistics), a statistical framework that quantifies the importance of genetic variants to a phenotype based on each individuals ancestry background, and profiles heterogeneous genetic effects across the genetic ancestry continuum. We demonstrate the utility of this framework through both simulations and real data analysis using sequencing data from ancestrally diverse cohorts in the Alzheimers Disease Sequencing Project (ADSP). We show that Alzheimers Disease (AD) risk variants span a spectrum from ancestry-homogeneous to ancestry-dependent effects, and that PFstatistics characterizes this spectrum at individual resolution across the ancestry continuum. PFstatistics also provides individual-level variant selection with FDR controlled at a target level, yielding distinct selection sets that vary across individuals according to their ancestry background. While demonstrated in the context of genetic ancestry, the proposed method is broadly applicable to other heterogeneity features such as environmental factors, offering a robust tool for understanding complex genetic contributions across diverse populations.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Combining case-control status and family history of disease increases association power 96%
- Improving Polygenic Prediction in Ancestrally Diverse Populations 96%
- Leveraging functional genomic annotations and genome coverage to improve polygenic prediction of complex traits within and between ancestries 96%
Similar papers in this journal
- Identification of putative causal loci in whole-genome sequencing data via knockoff statistics 96%
- Quantifying portable genetic effects and improving cross-ancestry genetic prediction with GWAS summary statistics 96%
- Population-specific causal disease effect sizes in functionally important regions impacted by selection 96%
Similar papers in this journal
- A statistical framework for powerful multi-trait rare variant analysis in large-scale whole-genome sequencing studies 94%
- Exploiting pleiotropy to enhance variant discovery with functional false discovery rates 94%
- MetaSTAARlite: An all-in-one tool for biobank-scale whole-genome sequencing meta-analysis 93%
Similar papers in this journal
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.