DeepEXPOKE: A Deep Learning Framework with Polygenic Risk Scores as Knockoffs for Deconvoluting Genetic and Non-Genetic Exposure Risks in Sepsis and Coronary Heart Disease
Sriram, A.; Park, H. J.; Carcillo, J. A.; Kernan, K. F.; Bohn, R. C.; Kim, S.
Show abstract
The exposome refers to the totality of environmental, behavioral, and lifestyle exposures an individual experiences throughout ones lifetime. Due to the modifiability of exposures, identifying the risk exposures on a disease is crucial for effective intervention and prevention of the disease. However, traditional analytical methods struggle to capture the complexities of exposome data: nonlinear effects, correlated exposures, and potential interplay with genetic effects. To address these challenges and accurately estimate exposure effects on complex diseases, we developed DeepEXPOKE, a deep learning framework integrating two types of knockoff features: statistical knockoffs (statKO) and polygenic risk score as knockoffs (PRSKO). DeepEXPOKE-statKO controls exposure correlation and DeepEXPOKE-PRSKO isolates genetic effects, while both can capture nonlinear effects. We applied DeepEXPOKE to predict outcomes of two significant diseases with distinct etiology and clinical presentation: sepsis and coronary heart disease (CHD), demonstrating its performance in comparison to existing machine learning methods. Furthermore, both DeepEXPOKE-PRSKO and DeepEXPOKE-statKO identified metabolites such as glucose and triglycerides as risk factors for sepsis and suggested that their effects are primarily at the non-genetic level, consistent with the role of metabolites in responding to environmental factors. Additionally, DeepEXPOKE-PRSKO uniquely identified asthma as a sepsis risk factor and suggested its effect is partially at the genetic level, offering insights into the conflicting associations observed between the genome data studies and patient data analysis regarding asthma and sepsis risk. Overall, DeepEXPOKE offers a novel DNN approach for identifying and interpreting exposure risk factors, advancing our understanding of complex diseases.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Genetic analysis of blood molecular phenotypes reveals regulatory networks affecting complex traits: a DIRECT study 96%
- OTTERS: A powerful TWAS framework leveraging summary-level reference data 95%
- Deep representation learning for clustering longitudinal survival data from electronic health records 95%
Similar papers in this journal
Similar papers in this journal
- Causal effects of maternal circulating amino acids on offspring birthweight: a Mendelian randomisation study 94%
- Multi-ancestry omic Mendelian randomization revealing putative drug targets of COVID-19 severity 93%
- Integrative deep learning analysis improves colon adenocarcinoma patient stratification at risk for mortality 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.