Ubie Symptom Checker: A Clinical Vignette Simulation Study
Taylor, N. K.; Nishibayashi, T.
Show abstract
BackgroundAI-driven symptom checkers (SC) are increasingly adopted in healthcare for their potential to provide users with accessible and immediate preliminary health education. These tools, powered by advanced artificial intelligence algorithms, assist patients in quickly assessing their symptoms. Previous studies using clinical vignette approaches have evaluated SC accuracy, highlighting both strengths and areas for improvement. ObjectiveThis study aims to evaluate the performance of the Ubie Symptom Checker (Ubie SC) using an innovative large language model-assisted (LLM) simulation method. MethodsThe study employed a three-phase methodology: gathering 400 publicly available clinical vignettes, medical entity linking these vignettes to the Ubie SC using large language models and physician supervision, and evaluation of accuracy metrics. The analysis focused on 328 vignettes that were within the scope of the Ubie SC with accuracy measured by Top-5 hit rates. ResultsUbie achieved a Top-5 hit accuracy of 63.4% and a Top-10 hit accuracy of 71.6%, indicating its effectiveness in providing relevant information based on symptom input. The system performed particularly well in domains such as the nervous system and respiratory conditions, though variability in accuracy was observed across different ICD groupings, highlighting areas for further refinement. When compared to physicians and comparator SCs that used the same clinical vignettes set, Ubie compared favorably to the median physician hit accuracy. ConclusionsThe Ubie Symptom Checker shows considerable promise as a supportive education tool in healthcare. While the study highlights the systems strengths, it also identifies areas for improvement suggesting continued refinement and real-world testing are essential to fully realize Ubies potential in AI-assisted healthcare.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Development of a customised data management system for a COVID-19-adapted colorectal cancer pathway 95%
- User Testing of a Diagnostic Decision Support System with Machine-assisted Chart Review to Facilitate Clinical Genomic Diagnosis 94%
- Connecting Artificial Intelligence and Primary Care Challenges: Findings from a Multi-Stakeholder Collaborative Consultation 94%
Similar papers in this journal
- Harnessing the Open Access Version of ChatGPT for Enhanced Clinical Opinions 96%
- Theory of radiologist interaction with instant messaging decision support tools: a sequential-explanatory study 95%
- Cardiology Knowledge Assessment of Retrieval-Augmented Open versus Proprietary Large Language Models 94%
Similar papers in this journal
- Development and Validation of ‘Patient Optimizer’ (POP) Algorithms for Predicting Surgical Risk with Machine Learning 94%
- On the predictability of postoperative complications for cancer patients: a Portuguese cohort study 92%
- Evaluating Semantic Similarity Methods for Comparison of Text-derived Phenotype Profiles 92%
Similar papers in this journal
- The potential for digital patient symptom recording through symptom assessment applications to optimize patient flow and reduce waiting times in Urgent Care Centers: a simulation study 95%
- Improving emergency department patient-doctor conversation through an artificial intelligence symptom taking tool: an action-oriented design pilot study 94%
- Is virtual care the new normal? Evidence supporting Covid-19’s durable transformation on healthcare delivery 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.