Incorporating Dietary Information to Enhance Polygenic Prediction Models with Applications to Body Mass Index and Type 2 Diabetes
LEE, E. Y.; Dinh, B. L.; Tang, J.; Streicher, S.; Biswas, S.; Tian, H.; Maskarinec, G.; Wilkens, L. R.; Park, S.-Y.; Chiang, C. W. K.
Show abstract
Polygenic predictors can enhance screening for biomedical conditions, such as metabolism-related traits and diseases, but explain limited phenotypic variance and face implementation challenges in non-European populations. On the other hand, dietary quality and other sociocultural factors are well established metabolic risk factors that remain under-investigated in risk stratification models. In this study, we developed and evaluated risk stratification model combining polygenic predictors and diet-based models for body mass index (BMI) and type 2 diabetes (T2D). Using 5,368 Native Hawaiians from the Multiethnic Cohort (MEC-NH) with genetic data, we integrated large-scale cross-ancestry GWAS summary statistics to develop polygenic score (PGS) models with better prediction accuracies (partial-R2 [SE] = 0.12 [0.04] for BMI; liability-R2 [SE] = 0.09 [0.04] and AUC = 0.65 for T2D) than using GWAS information from single ancestry (partial-R2 = 0.03-0.09 for BMI and liability-R2 = 0.01-0.07 and AUC = 0.52-0.63 for T2D) or in combination with GWAS from MEC-NH (partial-R2 = 0.04-0.07 for BMI and liability-R2 = 0.01-0.04 and AUC = 0.53-0.62). Moreover, machine learning models trained on 520 dietary variables available from 14,344 MEC-NH individuals substantially explained BMI variation (partial-R2 [SE] = 0.12 [0.01]) and enhanced prediction when combined with PGS (adjusted-R2 = 0.29). The best performing diet score model for BMI was associated with multiple chronic diseases in the same cohort, potentially mediated via inflammatory and lipid pathways. Both PGS and dietary scores provided significant, complementary information for predicting BMI and T2D. For populations with limited genetic studies like Native Hawaiians, integrating external GWAS data or non-genetic dietary information can improve risk stratification models.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Genome-wide association study of body fat distribution traits in Hispanics/Latinos from the HCHS/SOL Study 96%
- Genome-wide gene-diet interaction analysis in the UK Biobank identifies novel effects on Hemoglobin A1c 95%
- Evaluating and implementing block jackknife resampling Mendelian randomization to mitigate bias induced by overlapping samples 95%
Similar papers in this journal
- Plasma proteomic signatures of adiposity are associated with cardiovascular risk factors and type 2 diabetes risk in a multi-ethnic Asian population 96%
- Genome-Wide Association Meta-Analysis Using a Recessive Model Illuminates Genetic Architecture of Type 2 Diabetes 95%
- Characterizing common and rare variations in non-traditional glycemic biomarkers using multivariate approaches on multi-ancestry ARIC study 94%
Similar papers in this journal
- PhenomeXcan: Mapping the genome to the phenome through the transcriptome 94%
- Sex-specific phenotypic effects and evolutionary history of an ancient polymorphic deletion of the human growth hormone receptor 94%
- Nicotinamide riboside improves muscle mitochondrial biogenesis, satellite cell differentiation and gut microbiota composition in a twin study 93%
Similar papers in this journal
- Genome-wide study on 72,298 Korean individuals in Korean biobank data for 76 traits identifies hundreds of novel loci 93%
- Taiwan Biobank: a rich biomedical research database of the Taiwanese population 93%
- Integrative polygenic risk score improves the prediction accuracy of complex traits and diseases 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.