An explainable machine learning-based phenomapping strategy for adaptive predictive enrichment in randomized controlled trials
Oikonomou, E. K.; Thangaraj, P. M.; Bhatt, D. L.; Ross, J. S.; Young, L. H.; Krumholz, H. M.; Suchard, M. A.; Khera, R.
Show abstract
Randomized controlled trials (RCT) represent the cornerstone of evidence-based medicine but are resource-intensive. We propose and evaluate a machine learning (ML) strategy of adaptive predictive enrichment through computational trial phenomaps to optimize RCT enrollment. In simulated group sequential analyses of two large cardiovascular outcomes RCTs of (1) a therapeutic drug (pioglitazone versus placebo; Insulin Resistance Intervention after Stroke (IRIS) trial), and (2) a disease management strategy (intensive versus standard systolic blood pressure reduction in the Systolic Blood Pressure Intervention Trial (SPRINT)), we constructed dynamic phenotypic representations to infer response profiles during interim analyses and examined their association with study outcomes. Across three interim timepoints, our strategy learned dynamic phenotypic signatures predictive of individualized cardiovascular benefit. By conditioning a prospective candidates probability of enrollment on their predicted benefit, we estimate that our approach would have enabled a reduction in the final trial size across ten simulations (IRIS: - 14.8% {+/-} 3.1%, pone-sample t-test=0.001; SPRINT: -17.6% {+/-} 3.6%, pone-sample t-test<0.001), while preserving the original average treatment effect (IRIS: hazard ratio of 0.73 {+/-} 0.01 for pioglitazone vs placebo, vs 0.76 in the original trial; SPRINT: hazard ratio of 0.72 {+/-} 0.01 for intensive vs standard systolic blood pressure, vs 0.75 in the original trial; all with pone-sample t-test<0.01). This adaptive framework has the potential to maximize RCT enrollment efficiency.
Matching journals
The top 8 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- A Standardized Metric to Enhance Clinical Trial Design and Outcome Interpretation in Type 1 Diabetes 94%
- Biomarker panels for improved risk prediction and enhanced biological insights in patients with atrial fibrillation 93%
- Cancer patient survival can be accurately parameterized, revealing time-dependent therapeutic effects and doubling the precision of small trials 91%
Similar papers in this journal
- Federated Target Trial Emulation using Distributed Observational Data for Treatment Effect Estimation 95%
- Cohort Design and Natural Language Processing to Reduce Bias in Electronic Health Records Research: The Community Care Cohort Project 94%
- Novel clinical subphenotypes in COVID-19: derivation, validation, prediction, temporal patterns, and interaction with social determinants of health 92%
Similar papers in this journal
- Genome-wide polygenic score with APOL1 risk genotypes predicts chronic kidney disease across major continental ancestries 92%
- Polygenic score informed by genome-wide association studies of multiple ancestries and related traits improves risk prediction for coronary artery disease 91%
- Actionable druggable genome-wide Mendelian randomization identifies repurposing opportunities for COVID-19 91%
Similar papers in this journal
- Collaborative Large Language Models for Automated Data Extraction in Living Systematic Reviews 92%
- Personalizing renal replacement therapy initiation in the intensive care unit: a reinforcement learning-based strategy with external validation on the AKIKI randomized controlled trials 91%
- A Web-based Tool for Automatically linking Clinical Trials to their Publications 90%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.