Statistical refinement of case vignettes for digital health research
Kopka, M.; Feufel, M. A.
Show abstract
Digital health research often relies on case vignettes (descriptions of fictitious or real patients) to navigate ethical and practical challenges. Despite their utility, the quality and lack of standardization of these vignettes has often been criticized, especially in studies on symptom-assessment applications (SAAs) and triage decision-making. To address this, our paper introduces a method to refine an existing set of vignettes, drawing on principles from classical test theory. First, we removed any vignette with an item difficulty of zero and an item-total correlation below zero. Second, we stratified the remaining vignettes to reflect the natural base rates of symptoms that SAAs are typically approached with, selecting those vignettes with the highest item-total correlation in each quota. Although this two-step procedure reduced the size of the original vignette set by 40%, comparing triage performance on the reduced and the original vignette sets, we found a strong correlation (r = 0.747 to r = 0.997, p < .001). This indicates that using our refinement method helps identifying vignettes with high predictive power of an agents triage performance while simultaneously increasing cost-efficiency of vignette-based evaluation studies. This might ultimately lead to higher research quality and more reliable results.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Essential Indicators of Quality in Primary Care Settings: An Evidence-Based, Structured, Expert Approach 94%
- Artificial Intelligence for Contextual Well-being: Protocol for an Exploratory Sequential Mixed Methods Study with Medical Students as a Social Microcosm 93%
- Evaluating user experience with immersive technology in simulation-based education: a modified Delphi study with qualitative analysis 93%
Similar papers in this journal
Similar papers in this journal
- Assessing ChatGPT’s Mastery of Bloom’s Taxonomy using psychosomatic medicine exam questions 94%
- Health indicators as a measure of individual health status: public perspectives 92%
- Using a Multilingual AI Care Agent to Reduce Disparities in Colorectal Cancer Screening: Higher FIT Test Adoption Among Spanish-Speaking Patients 91%
Similar papers in this journal
- Pre-Post Analysis of the Impact of British Columbia Nurse Practitioner Primary Care Clinics on Patient Health and Care Experience 94%
- What is the suitability of clinical vignettes in benchmarking the performance of online symptom checkers? An audit study 94%
- A mixed methods study protocol to develop and pilot a Competency Assessment Tool to support therapists in the care of patients with blunt CHest trauma (CATCh Study) 92%
Similar papers in this journal
- Measures of socioeconomic advantage are not independent predictors of support for healthcare AI: subgroup analysis of a national Australian survey 92%
- Connecting Artificial Intelligence and Primary Care Challenges: Findings from a Multi-Stakeholder Collaborative Consultation 92%
- Development of a customised data management system for a COVID-19-adapted colorectal cancer pathway 91%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.