Diagnostic test accuracy of artificial intelligence in screening for referable diabetic retinopathy in real-world settings: A systematic review and meta-analysis
Uy, H.; Fielding, C.; Hohlfeld, A.; Ochodo, E.; Opare, A.; Mukonda, E.; Minnies, D.; Engel, M. E.
Show abstract
Studies on artificial intelligence (AI) in screening for diabetic retinopathy (DR) have shown promising results in addressing the mismatch between the capacity to implement DR screening and the increasing DR incidence; however, most of these studies were done retrospectively. This review sought to evaluate the diagnostic test accuracy (DTA) of AI in screening for referable diabetic retinopathy (RDR) in real-world settings. We searched CENTRAL, PubMed, CINAHL, Scopus, and Web of Science on 9 February 2023. We included prospective DTA studies assessing AI against trained human graders (HGs) in screening for RDR in patients living with diabetes. synthesis Two reviewers independently extracted data and assessed methodological quality against QUADAS-2 criteria. We used the hierarchical summary receiver operating characteristics (HSROC) model to pool estimates of sensitivity and specificity and, forest plots and SROC plots to visually examine heterogeneity in accuracy estimates. Finally, we conducted sensitivity analyses to explore the effects of studies deemed to possibly affect the quality of the studies. We included 15 studies (17 datasets: 10 patient-level analysis (N=45,785), and 7 eye-level analysis (N=15,390). Meta-analyses revealed a pooled sensitivity of 95.33%(95% CI: 90.60-100%) and specificity of 92.01%(95% CI: 87.61-96.42%) for patient-level analysis; for the eye-level analysis, pooled sensitivity was 91.24% (95% CI: 79.15-100%) and specificity, 93.90% (95% CI: 90.63-97.16%). Subgroup analyses did not provide variations in the diagnostic accuracy of country classification and DR classification criteria; however, a moderate increase was observed in diagnostic accuracy at the primary-level and, a minimal decrease in the tertiary-level healthcare settings. Sensitivity analyses did not show any variations in studies that included diabetic macular edema in the RDR definition, nor in studies with [≥]3 HGs. This review provides evidence, for the first time from prospective studies, for the effectiveness of AI in screening for RDR, in real-world settings.
Matching journals
The top 2 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Identifying the Influencing Factors for Cataract Surgery Uptake in Malaysia 95%
- The evaluation of a web‐based tool for measuring the uncorrected visual acuity and refractive error in keratoconus eyes: a prospective open‐label method comparison study 95%
- Development of the Advised Protocol for OCT Study Terminology and Elements Anterior Segment OCT extension reporting guidelines: APOSTEL-AS 95%
Similar papers in this journal
- Diagnostic Performance of Deep Learning in Infectious Keratitis: A Systematic Review and Meta-Analysis Protocol 96%
- Prevalence, incidence and risk factors for myopia among urban and rural children in southern China: protocol for a school-based cohort study 94%
- One and Two Year Visual Outcomes from the Moorfields AMD Database - an Open Science Resource for the Study of Neovascular Age-related Macular Degeneration 94%
Similar papers in this journal
- Risk Model for Intraoperative Complication during Cataract Surgery Based on Data from 900,000 Eyes – Previous Intravitreal Injection is a Risk Factor 95%
- Unveiling the Clinical Incapabilities: A Benchmarking Study of GPT-4V(ision) for Ophthalmic Multimodal Image Analysis 94%
- The Moorfields AMD Database Report 2 - Fellow Eye Involvement with Neovascular Age-related Macular Degeneration 94%
Similar papers in this journal
- Increasing frequency of hospital admissions for retinal detachment and vitreo-retinal surgery in England 2000-2018 94%
- An Open-Source Dataset Of Anti-Vegf Therapy In Diabetic Macular Oedema Patients Over Four Years & Their Visual Outcomes 93%
- Evaluation of OCT biomarker changes in treatment-naive neovascular AMD using a deep semantic segmentation algorithm 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.