An LLM-assisted framework for accelerated and verifiable clinical hypothesis testing from electronic health records
Gim, N.; Gim, I.; Jiang, Y.; Kihara, Y.; Blazes, M.; Wu, Y.; Lee, C. S.; Lee, A. Y.
Show abstract
Acquiring insights from electronic health records (EHRs) is slowed by manual analytical workflows that limit scalability and reproducibility. We present LATCH (LLM-Assisted Testing of Clinical Hypotheses), an agentic framework that converts natural language clinical hypotheses into fully auditable analyses on structured EHR data. LATCH integrates LLM-assisted semantic layers with deterministic execution pipelines to automate cohort construction, statistical analysis, and result reporting, while isolating patient-level data from LLM-involved steps. Using diabetes as a model disease, LATCH reproduced findings from 20 published studies within 3-15 minutes per study. Beyond replication, LATCH enabled study extensions and new insight generation through simple natural language hypothesis modifications. We demonstrated LATCH across 102 hypothesis tests spanning reproduction, extension, and insight generation. We systematically stress-tested LATCH to characterize its limitations and operational boundaries. LATCH provides a scalable framework for reproducible real-world evidence generation, reducing analytical bottlenecks and improving reliability of AI-assisted biomedical discovery while preserving human oversight.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Clinical Knowledge Extraction via Sparse Embedding Regression (KESER) with Multi-Center Large Scale Electronic Health Record Data 93%
- Finding Long-COVID: Temporal Topic Modeling of Electronic Health Records from the N3C and RECOVER Programs 93%
- Federated Target Trial Emulation using Distributed Observational Data for Treatment Effect Estimation 93%
Similar papers in this journal
- The Interpretable Multimodal Machine Learning (IMML) framework reveals pathological signatures of distal sensorimotor polyneuropathy 93%
- Pretrained Patient Trajectories for Adverse Drug Event Prediction Using Common Data Model-based Electronic Health Records 93%
- A user-friendly tool for cloud-based whole slide image segmentation, with examples from renal histopathology 92%
Similar papers in this journal
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.