Impact of an Artificial Intelligence Algorithm on Diabetic Retinopathy Grading by Ophthalmology Residents
Paul, S. K.; Kim, C. U.; Shieh, D.; Zhou, X. Y.; Pan, I.; Mehra, A. A.; Sobol, W.
Show abstract
PurposeTo determine whether AI significantly affects the performance of diabetic retinopathy (DR) grading by ophthalmology residents. Secondary objectives included evaluation of AIs effects on intergrader variability, self-reported confidence, and decision making. MethodsFour ophthalmology residents at a single academic medical center across all years of training (PGY-2 to PGY-4) analyzed 265 retinal fundus photographs for diabetic retinopathy from a publicly available dataset without and with the assistance of an AI algorithm, separated by a 3-week washout period ResultsOverall, there was no significant difference without versus with AI in five-class grading, as measured by QWK, with differences ranging from +0.010-0.017, p=0.09-0.32. No significant difference without and with AI was observed for binary classification of referable DR, except for the specificity of the PGY-3 resident (71.8% to 80%, p=0.019). Intergrader agreement among residents significantly increased with AI (FK +0.072, p=0.0003). Self-reported confidence also significantly increased for 3 out of 4 residents. ConclusionThe use of an AI algorithm did not significantly affect the DR grading performance of ophthalmology residents but did increase intergrader agreement and self-reported confidence. Introducing AI into the ophthalmology residency curriculum may be beneficial as the technology becomes more prevalent. Summary StatementA cross-sectional study that evaluated the performance of ophthalmology residents grading diabetic retinopathy fundus photographs with and without the assistance of an artificial intelligence algorithm.
Matching journals
The top 2 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- The learning curve of murine subretinal injection among clinically trained ophthalmic surgeons 96%
- Visual field evaluation using Zippy Adaptive Threshold Algorithm (ZATA) Standard and ZATA Fast in patients with glaucoma and healthy individuals 95%
- Relating Standardized Automated Perimetry Performed with Stimulus Sizes III and V in Eyes With Field Loss due to Glaucoma and NAION 95%
Similar papers in this journal
- Unveiling the Clinical Incapabilities: A Benchmarking Study of GPT-4V(ision) for Ophthalmic Multimodal Image Analysis 95%
- Automated Expert-level Scleral Spur Detection and Quantitative Biometric Analysis on the ANTERION Anterior Segment OCT System 95%
- Macula structural and vascular differences in glaucoma eyes with and without high axial myopia 94%
Similar papers in this journal
- Automated vision screening of children using a mobile graphic device 95%
- Evaluation of OCT biomarker changes in treatment-naive neovascular AMD using a deep semantic segmentation algorithm 93%
- Increasing frequency of hospital admissions for retinal detachment and vitreo-retinal surgery in England 2000-2018 93%
Similar papers in this journal
- Towards implementation of AI in New Zealand national screening program: Cloud-based, Robust, and Bespoke 94%
- Ocular findings, surgery details and outcomes in proliferative diabetic retinopathy patients with chronic kidney disease 94%
- The evaluation of a web‐based tool for measuring the uncorrected visual acuity and refractive error in keratoconus eyes: a prospective open‐label method comparison study 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.