Integrating genetic data in target trial emulations improves their design and informs the value of polygenic scores for prognostic and predictive enrichment
German, J.; Yang, Z.; Urbut, S.; Vartianinen, P.; FinnGen, ; Natarajan, P.; Pattorno, E.; Kutalik, Z.; Philippakis, A.; Ganna, A.
Show abstract
Randomized controlled trials (RCTs) are the gold standard for evaluating the efficacy and safety of medical interventions but ethical, practical, and financial limitations often necessitate decisions based on observational data. The increasing volume of such data has prompted regulatory bodies to rely more on real-world evidence, primarily obtained through trial emulations. This study explores how genetic data can improve the design of both emulated and traditional trials. We successfully emulated four major cardiometabolic RCTs within FinnGen (N=425 483) and showed how reduced differences in polygenic scores (PGS) between trial arms track improved study design and consequently reduced residual confounding. Complementing these results with simulations, we show that PGS cannot be directly used to adjust for residual or unmeasured confounding. Instead, we propose an approach that uses genetic instruments for confounding detection and apply this approach to identify likely confounders in Empareg trial emulation. Finally, our results suggest that trial emulations can inform the practical application of PGS in RCTs, potentially improving statistical power. Such prognostic enrichment strategies need to be assessed in a trial-relevant population, and we show that, for 2 out of 4 emulated trials, the association between PGS and trial outcomes in the general population was different from what observed in the population included in the trial. In conclusion, our work shows that genetic information can improve the design of emulated trials. These results contribute to the establishment of a promising new era of genetically-informed clinical trials.
Matching journals
The top 8 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Within-family studies for Mendelian randomization: avoiding dynastic, assortative mating, and population stratification biases 94%
- A Standardized Metric to Enhance Clinical Trial Design and Outcome Interpretation in Type 1 Diabetes 94%
- A novel Mendelian randomization method identifies causal relationships between gene expression and low-density lipoprotein cholesterol levels. 93%
Similar papers in this journal
- Quantifying absolute treatment effect heterogeneity for time-to-event outcomes across different risk strata: divergence of conclusions with risk difference and restricted mean survival difference 94%
- A phenome-wide multi-directional Mendelian randomization analysis of atrial fibrillation 93%
- High-throughput multivariable Mendelian randomization analysis prioritizes apolipoprotein B as key lipid risk factor for coronary artery disease 92%
Similar papers in this journal
- The Proportion of Randomized Controlled Trials That Inform Clinical Practice: A Longitudinal Cohort Study of Trials Registered on ClinicalTrials.gov 91%
- Phenome-wide Mendelian randomization study of plasma triglycerides and 2,600 disease traits 91%
- Sparse Dimensionality Reduction Approaches in Mendelian Randomization with highly correlated exposures 91%
Similar papers in this journal
- Federated Target Trial Emulation using Distributed Observational Data for Treatment Effect Estimation 95%
- Cohort Design and Natural Language Processing to Reduce Bias in Electronic Health Records Research: The Community Care Cohort Project 95%
- Novel clinical subphenotypes in COVID-19: derivation, validation, prediction, temporal patterns, and interaction with social determinants of health 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.