A versatile, fast and unbiased method for estimation of gene-by-environment interaction effects on biobank-scale datasets
Khan, M. A.; Di Scipio, M.; Judge, C.; Perrot, N.; Chong, M.; Mao, S.; Di, S.; Nelson, W.; Petch, J.; Pare, G.
Show abstract
Current methods to evaluate gene-by-environment (GxE) interactions on biobank-scale datasets are limited. MonsterLM enables multiple linear regression on genome-wide datasets, does not rely on parameters specification and provides unbiased estimates of variance explained by GxE interaction effects. We applied MonsterLM to the UK Biobank for eight blood biomarkers (N=325,991), identifying significant genome-wide interaction variance with waist-to-hip ratio for five biomarkers, with variance explained by interactions ranging from 0.11 to 0.58. 48% to 94% of GxE interaction variance can be attributed to variants without significant marginal association with the phenotype of interest. Conversely, for most traits, >40% of interaction variance was explained by less than 5% of genetic variants. We observed significant improvements in polygenic score prediction with incorporation of GxE interactions in four biomarkers. Our results imply an important contribution of GxE interaction effects, driven largely by a restricted set of variants distinct from loci with strong marginal effects.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- A novel Mendelian randomization method identifies causal relationships between gene expression and low-density lipoprotein cholesterol levels. 97%
- A novel method for an unbiased estimate of cross-ancestry genetic correlation using individual-level data 96%
- Testing and controlling for horizontal pleiotropy with the probabilistic Mendelian randomization in transcriptome-wide association studies 96%
Similar papers in this journal
- Pathway-specific polygenic scores substantially increase the discovery of gene-adiposity interactions impacting liver biomarkers 96%
- Polygenic risk score prediction accuracy convergence 95%
- Multivariate adaptive shrinkage improves cross-population transcriptome prediction for transcriptome-wide association studies in underrepresented populations 95%
Similar papers in this journal
- Composite trait Mendelian Randomization reveals distinct metabolic and lifestyle consequences of differences in body shape 96%
- Capturing additional genetic risk from family history for improved polygenic risk prediction 95%
- Polygenic Risk Prediction using Gradient Boosted Trees Captures Non-Linear Genetic Effects and Allele Interactions in Complex Phenotypes 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.