Identifying and Characterizing Gallstone Disease from Clinical Narratives with Zero-shot Learning and Automated Prompt Optimization
Hwang, S.; Wang, A.; Batugo, A.; Kaplan, D. E.; Rader, D.; Mowery, D.; Lim, J.
Show abstract
We built and evaluated a zero-shot LLM pipeline with automated, task-aware prompt optimization to extract radiology and symptom fields for gallstone phenotyping from de-identified EHR text. Across symptomatic, asymptomatic, and control cohorts, it performed reliably on high-signal binary fields and symptom flags but lagged on fine-grained stone burden and complications, establishing a practical baseline and motivating targeted refinements
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Natural language inference for clinical registry curation 92%
- Large Language Models Facilitate the Generation of Electronic Health Record Phenotyping Algorithms 92%
- Development and Validation of Phenotype Classifiers across Multiple Sites in the Observational Health Sciences and Informatics (OHDSI) Network 92%
Similar papers in this journal
- Modular Clinical Decision Support Networks (MoDN)—Updatable, Interpretable, and Portable Predictions for Evolving Clinical Environments 91%
- Development and Validation of a Deep Learning Model for Detecting Signs of Tuberculosis on Chest Radiographs among US-bound Immigrants and Refugees 91%
- Performance of Generative Pretrained Transformer on the National Medical Licensing Examination in Japan 90%
Similar papers in this journal
- Design and Implementation of an End-to-End AI-Driven Colonoscopy Recall Workflow at Scale 92%
- Large-Scale Deep Learning for Metastasis Detection in Pathology Reports 91%
- MMFP-Tableau: Enabling Precision Mitochondrial Medicine through Integration, Visualization, and Analytics of Clinical and Research Health System Electronic Data 91%
Similar papers in this journal
- Towards a Clinically-based Common Coordinate Framework for the Human Gut Cell Atlas - The Gut Models 92%
- ARDSFlag: An NLP/Machine Learning Algorithm to Visualize and Detect High-Probability ARDS Admissions Independent of Provider Recognition and Billing Codes 91%
- Temporal Relationship of Computed and Structured Diagnoses in Electronic Health Record Data 90%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.