An integrative analysis of genomic and exposomic data for complex traits and phenotypic prediction
Zhou, X.; Lee, S. H.
Show abstract
Complementary to the genome, the concept of exposome has been proposed to capture the totality of human environmental exposures. While there has been some recent progress on the construction of the exposome, few tools exist that can integrate the genome and exposome for complex trait analyses. Here we propose a linear mixed model approach to bridge this gap, which jointly models the random effects of the two omics layers on phenotypes of complex traits. We illustrate our approach using traits from the UK Biobank (e.g., BMI & height for N [~] 35,000) with a small fraction of the exposome that comprises 28 lifestyle factors. The joint model of the genome and exposome explains substantially more phenotypic variance and significantly improves phenotypic prediction accuracy, compared to the model based on the genome alone. The additional phenotypic variance captured by the exposome includes its additive effects as well as non-additive effects such as genome-exposome (gxe) and exposome-exposome (exe) interactions. For example, 19% of variation in BMI is explained by additive effects of the genome, while additional 7.2% by additive effects of the exposome, 1.9% by exe interactions and 4.5% by gxe interactions. Correspondingly, the prediction accuracy for BMI, computed using Pearsons correlation between the observed and predicted phenotypes, improves from 0.15 (based on the genome alone) to 0.35 (based on the genome & exposome). We also show, using established theories, integrating genomic and exposomic data is essential to attaining a clinically meaningful level of prediction accuracy for disease traits. In conclusion, the genomic and exposomic effects can contribute to phenotypic variation via their latent relationships, i.e. genome-exposome correlation, and gxe and exe interactions, and modelling these effects has a great potential to improve phenotypic prediction accuracy and thus holds a great promise for future clinical practice.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- A novel method for an unbiased estimate of cross-ancestry genetic correlation using individual-level data 96%
- A novel Mendelian randomization method identifies causal relationships between gene expression and low-density lipoprotein cholesterol levels. 96%
- Testing and controlling for horizontal pleiotropy with the probabilistic Mendelian randomization in transcriptome-wide association studies 95%
Similar papers in this journal
- Controlling for background genetic effects using polygenic scores improves the power of genome-wide association studies 95%
- Quantifying genetic heterogeneity between continental populations for human height and body mass index 94%
- The Impact of Late-Career Job Loss and Genotype on Body Mass Index 94%
Similar papers in this journal
- GxE PRS: Genotype-environment interaction in polygenic risk score models for quantitative and binary traits 95%
- Identity-by-descent mapping using multi-individual IBD with genome-wide multiple testing adjustment 94%
- Efficient gene-environment interaction tests for large biobank-scale sequencing studies 92%
Similar papers in this journal
- Composite trait Mendelian Randomization reveals distinct metabolic and lifestyle consequences of differences in body shape 96%
- Capturing additional genetic risk from family history for improved polygenic risk prediction 95%
- Polygenic Risk Prediction using Gradient Boosted Trees Captures Non-Linear Genetic Effects and Allele Interactions in Complex Phenotypes 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.