Reassessing Instrument Strength in Two-Sample Mendelian Randomization Analysis
Liu, X.; Huang, Y.-J.; Purushotham, Y.; Sofer, T.
Show abstract
Mendelian randomization (MR) analysis is widely used to estimate causal relationships between risk factors and outcomes of interest. Two-sample MR approaches have gained increasing attention in genetic epidemiology due to the growing availability of Genome-Wide Association Study (GWAS) summary statistics from public databases. A critical step in two-sample MR is the selection of genetic variants as instrumental variables (IVs). Although genome-wide significant variants are typically preferred, the inclusion of variants with weaker association p-values is considered, as they may potentially improve power through an increased instrument number of instruments, while they may introduce weak instrument bias and attenuate effect estimates towards the null. Our simulation results show that even modest levels of pleiotropy substantially increase the variability of causal effect estimates, while the inclusion of weak IVs does not substantially affect the direction and variability of causal effect estimates in most cases. In real data analyses, we used two released versions of FinnGen GWAS summary statistics with different sample sizes as exposure GWASs to assess the influence of weak IVs. Here, the inclusion of IVs with higher exposure-association p-values resulted in weakened estimated effect sizes, particularly when the exposure GWAS sample size was small. These findings suggest that incorporating weak IVs is reasonable when the exposure GWAS sample size is large, but it poses a risk of falsely concluding null associations when the exposure GWAS sample size is small.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Benchmarking Mendelian Randomization methods for causal inference using genome-wide association study summary statistics 96%
- A Bayesian approach to Mendelian randomization using summary statistics in the univariable and multivariable settings with correlated pleiotropy 95%
- A novel and efficient machine learning Mendelian randomization estimator applied to predict the safety and efficacy of sclerostin inhibition 95%
Similar papers in this journal
- Rare variants association testing for a binary outcome when pooling individual level data from heterogeneous studies 94%
- GxE PRS: Genotype-environment interaction in polygenic risk score models for quantitative and binary traits 93%
- A Two-stage Linear Mixed Model (TS-LMM) for Summary-data-based Multivariable Mendelian Randomization 93%
Similar papers in this journal
- Bias in two-sample Mendelian randomization when using heritable covariable-adjusted summary associations 96%
- A Comprehensive Evaluation of Methods for Mendelian Randomization Using Realistic Simulations and an Analysis of 38 Biomarkers for Risk of Type-2 Diabetes 95%
- An empirical investigation into the impact of winner's curse on estimates from Mendelian randomization 94%
Similar papers in this journal
- Accounting for genetic effect heterogeneity in fine-mapping and improving power to detect gene-environment interactions with SharePro 95%
- Simultaneous estimation of bi-directional causal effects and heritable confounding from GWAS summary statistics 94%
- A novel Mendelian randomization method identifies causal relationships between gene expression and low-density lipoprotein cholesterol levels. 94%
Similar papers in this journal
- Summary statistics from large-scale gene-environment interaction studies for re-analysis and meta-analysis 94%
- MR Corge: Sensitivity analysis of Mendelian randomization based on the core gene hypothesis for polygenic exposures 94%
- The predictive capacity of polygenic risk scores for disease risk is only moderately influenced by imputation panels tailored to the target population 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.