Expert-Level Detection of Referable Glaucoma from Fundus Photographs in a Safety Net Population: The AI and Teleophthalmology in Los Angeles Initiative
Nguyen, V.; Iyengar, S.; Rasheed, H.; Apolo, G.; Li, Z.; Kumar, A.; Nguyen, H.; Bohner, A.; Dhodapkar, R.; Do, J.; Duong, A.; Gluckstein, J.; Hong, K.; James, A.; Lee, J.; Nguyen, K.; Wong, B.; Ambite, J.-L.; Kesselman, C.; Daskivich, L.; Pazzani, M.; Xu, B.
Show abstract
PurposeTo develop and test a deep learning (DL) algorithm for detecting referable glaucoma in the Los Angeles County (LAC) Department of Health Services (DHS) teleretinal screening program. MethodsFundus photographs and patient-level labels of referable glaucoma (defined as cup-to-disc ratio [CDR] [≥] 0.6) provided by 21 trained optometrist graders were obtained from the LAC DHS teleretinal screening program. A DL algorithm based on the VGG-19 architecture was trained using patient-level labels generalized to images from both eyes. Area under the receiver operating curve (AUC), sensitivity, and specificity were calculated to assess algorithm performance using an independent test set that was also graded by 13 clinicians with one to 15 years of experience. Algorithm performance was tested using reference labels provided by either LAC DHS optometrists or an expert panel of 3 glaucoma specialists. Results12,098 images from 5,616 patients (2,086 referable glaucoma, 3,530 non-glaucoma) were used to train the DL algorithm. In this dataset, mean age was 56.8 {+/-} 10.5 years with 54.8% females and 68.2% Latinos, 8.9% Blacks, 2.7% Caucasians, and 6.0% Asians. 1,000 images from 500 patients (250 referable glaucoma, 250 non-glaucoma) with similar demographics (p [≥] 0.57) were used to test the DL algorithm. Algorithm performance matched or exceeded that of all independent clinician graders in detecting patient-level referable glaucoma based on LAC DHS optometrist (AUC = 0.92) or expert panel (AUC = 0.93) reference labels. Clinician grader sensitivity (range: 0.33-0.99) and specificity (range: 0.68-0.98) ranged widely and did not correlate with years of experience (p [≥] 0.49). Algorithm performance (AUC = 0.93) also matched or exceeded the sensitivity (range: 0.78-1.00) and specificity (range: 0.32-0.87) of 6 LAC DHS optometrists in the subsets of the test dataset they graded based on expert panel reference labels. ConclusionsA DL algorithm for detecting referable glaucoma developed using patient-level data provided by trained LAC DHS optometrists approximates or exceeds performance by ophthalmologists and optometrists, who exhibit variable sensitivity and specificity unrelated to experience level. Implementation of this algorithm in screening workflows could help reallocate eye care resources and provide more reproducible and timely glaucoma care.
Matching journals
The top 1 journal accounts for 50% of the predicted probability mass.
Similar papers in this journal
- Relating Standardized Automated Perimetry Performed with Stimulus Sizes III and V in Eyes With Field Loss due to Glaucoma and NAION 97%
- Rates of Glaucoma Progression Derived from Linear Mixed Models Using Varied Random Effect Distributions 97%
- Visual field evaluation using Zippy Adaptive Threshold Algorithm (ZATA) Standard and ZATA Fast in patients with glaucoma and healthy individuals 96%
Similar papers in this journal
- Automated Expert-level Scleral Spur Detection and Quantitative Biometric Analysis on the ANTERION Anterior Segment OCT System 97%
- Macula structural and vascular differences in glaucoma eyes with and without high axial myopia 96%
- Unveiling the Clinical Incapabilities: A Benchmarking Study of GPT-4V(ision) for Ophthalmic Multimodal Image Analysis 96%
Similar papers in this journal
- The evaluation of a web‐based tool for measuring the uncorrected visual acuity and refractive error in keratoconus eyes: a prospective open‐label method comparison study 96%
- Glaucoma Detection and Staging from Visual Field Images using Machine Learning Techniques 95%
- Prediction of the ectasia screening index from raw Casia2 volume data for keratoconus identification by using convolutional neural networks 95%
Similar papers in this journal
- Automated vision screening of children using a mobile graphic device 95%
- Evaluation of OCT biomarker changes in treatment-naive neovascular AMD using a deep semantic segmentation algorithm 95%
- Dense Optic Nerve Head Deformation Estimated using CNN as a Structural Biomarker of Glaucoma Progression 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.