Retrospective Evaluation of a Generative AI-Enabled Electronic Medical Record System in Primary Health Care Facilities in Kenya
Agweyu, A.; Mwaniki, P.; Musau, W.; Korom, R.; Isaaka, L.; Wanyama, C.; kiptinness, S.; Adan, N.; Emmanual-Fabula, M.; Mateen, B.
Show abstract
We conducted a retrospective evaluation of an electronic medical record-embedded large language model (LLM) clinical decision support system deployed across 16 primary care clinics in Kenya, between July-September 2024. A panel of trained physicians reviewed 1,469 records. Hallucinations were uncommon (50/1,469; 3.4%), most often involving mis-expanded acronyms or drug names. Clinical management guidance aligned with local guidelines in almost all cases (approximately 100%). Despite this, clinicians did not modify documentation in 62% of encounters. Safety assessments identified actively harmful recommendations from the LLM in 7.8% of encounters, with 67 such recommendations appearing in the final documentation. Conversely, risk present in the clinicians initial notes was fully mitigated in 118 encounters (8.0% overall; 12.1% of amended cases). Overall, the tool showed strong potential to support quality improvement, but the asymmetric adoption of harmful versus beneficial outputs underscore the need for usability optimization, local guardrails, and prospective trials to confirm patient-level benefit.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Self-tests for COVID-19: what is the evidence? A living systematic review and meta-analysis (2020-2023) 94%
- Implementing essential diagnostics-learning from essential medicines: A scoping review 94%
- Waiting times, patient flow, and occupancy density in South African primary health care clinics: implications for infection prevention and control 93%
Similar papers in this journal
Similar papers in this journal
- Automated and partially-automated contact tracing: a rapid systematic review to inform the control of COVID-19 94%
- Automated and semi-automated contact tracing: Protocol for a rapid review of available evidence and current challenges to inform the control of COVID-19 93%
- Remote Covid Assessment in Primary Care (RECAP) risk prediction tool: derivation and real-world validation studies 93%
Similar papers in this journal
- Connecting Artificial Intelligence and Primary Care Challenges: Findings from a Multi-Stakeholder Collaborative Consultation 94%
- User Testing of a Diagnostic Decision Support System with Machine-assisted Chart Review to Facilitate Clinical Genomic Diagnosis 93%
- The performance of national COVID-19 ‘Symptom Checkers’: A comparative case simulation study 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.