MRSL: A phenome-wide causal discovery algorithm based on GWAS summary data
Hou, L.; Geng, Z.; Shi, X.; Wang, C.; Li, H.; Xue, F.
Show abstract
Causal discovery is a powerful tool to disclose underlying structures by analyzing purely observational data. Genetic variants can provide useful complementary information for structure learning. Here, we propose a novel algorithm MRSL (Mendelian Randomization (MR)-based Structure Learning algorithm), which combines the graph theory with univariable and multivariable MR to learn the true structure using only GWAS summary statistics. Specifically, MRSL also utilizes topological sorting to improve the precision of structure learning and provides three adjusting categories for multivariable MR. Results of simulation reveal that MRSL has up to two-fold higher F1 score than other eight competitive methods. Additionally, the computing time of MRSL is 100 times faster than other methods. Furthermore, we apply MRSL to 26 biomarkers and 44 ICD10-defined diseases from UK Biobank. The results cover most of expected causal links which have biological interpretations and several new links supported by clinical case reports or previous observational literatures.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- CoxMDS: Multiple Data Splitting for High-dimensional Mediation Analysis with Survival Outcomes in Epigenome-wide Studies 96%
- HyMM: Hybrid method for disease-gene prediction by integrating multiscale module structures 95%
- BayesKAT: Bayesian Optimal Kernel-based Test for genetic association studies reveals joint genetic effects in complex diseases 95%
Similar papers in this journal
- Inferring Causal Direction Between Two Traits in the Presence of Horizontal Pleiotropy with GWAS Summary Data 96%
- Robust Inference of Bi-Directional Causal Relationships in Presence of Correlated Pleiotropy with GWAS Summary Data 96%
- A novel method for multiple phenotype association studies based on genotype and phenotype network 96%
Similar papers in this journal
- PAN: Personalized Annotation-based Networks for the Prediction of Breast Cancer Relapse 93%
- KGRACDA: A Model Based on Knowledge Graph from Recursion and Attention Aggregation for CircRNA-disease Association Prediction 93%
- Rapid Reconstruction of Time-varying Gene Regulatory Networks with Limited Main Memory 93%
Similar papers in this journal
- SynTL: A synthetic-data-based transfer learning approach for multi-center risk prediction 94%
- A Utility-Based Machine Learning-Driven Personalized Lifestyle Recommendation for Cardiovascular Disease Prevention 93%
- SurvMaximin: Robust Federated Approach to Transporting Survival Risk Prediction Models 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.