Leveraging a founder population to identify novel rare-population genetic determinants of lipidome
Montasser, M. E.; Aslibekyan, S.; Srinivasasainagendra, V.; Tiwari, H. K.; Patki, A.; Bagheri, M.; Kind, T.; Barupal, D.; Fan, S.; Perry, J. A.; Ryan, K. A.; Arnett, D. K.; Beitelshees, A. L.; Irvin, M. R.; O'Connell, J. R.
Show abstract
Identifying the genetic determinants of inter-individual variation in lipid species (lipidome) may provide deeper understanding and new insight into the mechanistic effect of complex lipidomic pathways in CVD risk and progression beyond simple traditional lipids. Previous studies have been largely population based and thus only powered to discover associations with common genetic variants. Founder populations represent a powerful resource to accelerate discovery of novel biology associated with rare population alleles that have risen to higher frequency due to genetic drift. We performed a GWAS of 355 lipid species in 650 individuals from the Old Order Amish founder population including 127 lipid species not previously tested. We report for the first time the lipid species associated with two rare-population but Amish-enriched lipid variants: APOB_rs5742904 and APOC3_rs76353203. We also identified novel associations for 3 rare-population Amish-enriched loci with several sphingolipids and with proposed potential functional/causal variant in each locus including GLPTD2_rs536055318, CERS5_rs771033566, and AKNA_rs531892793. We replicated 7 previously known common loci including novel associations with two sterols: androstenediol with UGT locus on chromosome 2 and estriol with SLC22A8/A24 locus on chromosome 11. Our results show the power of founder populations to discover novel biology due to genetic drift that can increase the frequency of an allele from only few copies in large sample cohorts such as the UK Biobank to dozens of copies in sample size as small as 650.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Rare variants in long non-coding RNAs are associated with blood lipid levels in the TOPMed Whole Genome Sequencing Study 97%
- A multi-layer functional genomic analysis to understand noncoding genetic variation in lipids 96%
- Widespread recessive effects on common diseases in a cohort of 44,000 British Pakistanis and Bangladeshis with high autozygosity 96%
Similar papers in this journal
- Whole genome sequence analysis of blood lipid levels in >66,000 individuals 98%
- Comprehensive genetic analysis of the human lipidome identifies novel loci controlling lipid homeostasis with links to coronary artery disease 97%
- Metabolome-wide Mendelian randomization characterizes heterogeneous and shared causal effects of metabolites on human health 96%
Similar papers in this journal
- Metabolic dysregulation of the lysophospholipid/autotaxin axis in the chromosome 9p21 gene SNP rs10757274 97%
- Metabolite Signature of Life’s Essential 8 and Risk of Coronary Heart Disease among Low-Income Black and White Americans 94%
- Coronary Artery Disease Risk of Familial Hypercholesterolemia Genetic Variants Independent of Historical Cholesterol Exposure 94%
Similar papers in this journal
- GWAS in Africans identifies novel lipids loci and demonstrates heterogenous association within Africa 97%
- Heritability and family-based GWAS analyses of the N-acyl ethanolamine and ceramide plasma lipidome 96%
- The impact of fatty acids biosynthesis on the risk of cardiovascular diseases in Europeans and East Asians: A Mendelian randomization study 95%
Similar papers in this journal
- Multi-trait genome-wide association study in 34,394 Chinese women reveals the genetic architecture of plasma metabolites during pregnancy 95%
- Genome-wide study on 72,298 Korean individuals in Korean biobank data for 76 traits identifies hundreds of novel loci 95%
- The genetic and phenotypic correlates of neonatal Complement Component 3 and 4 protein concentrations with a focus on psychiatric and autoimmune disorders 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.