Is it possible to vaccinate AI against bias? An exploratory study in epilepsy
Bhansali, R. M.; Westover, M. B.; Goldenholz, D. M.
Show abstract
ImportanceLarge language models are increasingly used for clinical decision support yet may perpetuate socioeconomic biases. Whether simple prompt-based interventions can mitigate such biases remains unknown. ObjectiveTo determine whether a prompt-based inoculation instructing large-language-models (LLMs) to disregard clinically irrelevant information can reduce bias and improve accuracy in recommendations. DesignExperimental study conducted November 21 to December 11, 2025. Each clinical vignette was presented 10 times per condition to account for stochastic variance. SettingPublicly available web interfaces of six frontier LLMs with memory features disabled. ParticipantsNo real patients were involved. Two fictional epilepsy vignettes (diagnostic and therapeutic) were created with identical clinical features but differing socioeconomic (SES) descriptors. Main Outcomes and MeasuresAccuracy (proportion of responses concordant with guidelines) and bias (accuracy difference between high and low SES vignettes), assessed via binary scoring based on evidence-based guidelines. ResultsA total of 480 LLM responses were analyzed. For diagnosis, base accuracy was 36% (43/120), with 45 percentage point bias gap (high SES 58% vs. low SES 13%); inoculation improved accuracy to 55% (66/120) and reduced bias to 27 percentage points. For treatment, base accuracy was 51% (61/120) with 25 percentage point bias gap; inoculation improved accuracy to 63% (75/120) and reduced bias to 8 percentage points. Responses to inoculation varied considerably: Gemini 3 Pro showed complete diagnostic bias elimination (low SES accuracy 0% [->] 100%), while Sonnet 4.5 showed paradoxical worsening. Conclusions and RelevanceA simple prompt-based intervention overall reduced socioeconomic bias and improved accuracy in LLM clinical recommendations, though effects varied across models. Prompt engineering may offer a practical approach to mitigating specific AI bias in healthcare. KEY POINTSO_ST_ABSQuestionC_ST_ABSCan a simple prompt-based "inoculation" instructing large language models to ignore clinically irrelevant socioeconomic details reduce bias and improve accuracy in epilepsy diagnosis and treatment recommendations? FindingsIn this experimental study of 480 responses from 6 large language models to paired high- vs low-socioeconomic status epilepsy vignettes, base diagnostic and treatment accuracies were 36% and 51%, respectively, with bias gaps of 45 and 25 percentage points, respectively; adding an inoculation prompt increased accuracy to 55% and 63% and reduced bias gaps to 27 and 8 percentage points, though effects varied by model, with some showing near-complete bias elimination and others demonstrating paradoxical worsening in certain conditions. MeaningPrompt-based inoculation may offer a practical, low-cost strategy to partially mitigate socioeconomic bias and modestly improve the quality of large language model clinical recommendations, but model-specific behavior and residual disparities highlight the need for ongoing oversight and complementary bias-mitigation strategies.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Evaluating the generalisability of region-naïve machine learning algorithms for the identification of epilepsy in low-resource settings 93%
- Development and preliminary testing of Health Equity Across the AI Lifecycle (HEAAL): A framework for healthcare delivery organizations to mitigate the risk of AI solutions worsening health inequities 91%
- Harnessing the Open Access Version of ChatGPT for Enhanced Clinical Opinions 91%
Similar papers in this journal
- Lived experiences of caregivers of persons with epilepsy attending an epilepsy clinic at a tertiary hospital, eastern Uganda: A phenomenological approach 92%
- Artificial Intelligence for Contextual Well-being: Protocol for an Exploratory Sequential Mixed Methods Study with Medical Students as a Social Microcosm 91%
- COVID-19 Control Strategies and Intervention Effects in Resource Limited Settings: A Modeling Study 91%
Similar papers in this journal
- Measures of socioeconomic advantage are not independent predictors of support for healthcare AI: subgroup analysis of a national Australian survey 92%
- Connecting Artificial Intelligence and Primary Care Challenges: Findings from a Multi-Stakeholder Collaborative Consultation 91%
- The performance of national COVID-19 ‘Symptom Checkers’: A comparative case simulation study 90%
Similar papers in this journal
- Empowering Personalized Pharmacogenomics with Generative AI Solutions 92%
- Use of unstructured text in prognostic clinical prediction models: a systematic review 91%
- Measuring Quality-of-Care in Treatment of Children with Attention-Deficit/Hyperactivity Disorder: A Novel Application of Natural Language Processing 90%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.