Enhancing Glaucoma Detection through Supervised Pre-training with Clinical Intermediate Indicators: A Multi-Cohort Deep Learning Approach
Li, Y.; Carrillo-Perez, F.; Alawad, M.; Gevaert, O.
Show abstract
Glaucoma is a leading cause of irreversible blindness worldwide, with early diagnosis often hindered by subtle symptomatology and the lack of comprehensive screening programs. In this study, we introduce a robust deep learning framework that leverages supervised pre-training with clinically relevant intermediate indicators, most notably the vertical cup-to-disc ratio (VCDR), to enhance glaucoma detection from color fundus images. Utilizing the expansive AIROGS dataset for pre-training, our multi-task learning strategy simultaneously addresses categorical diagnostic classification and VCDR regression. We evaluated three architectures, ResNet-18, DINOv2, and RETFound across multiple international cohorts, including DRISHTI, G1020, ORIGA, PAPILA, REFUGE1, and ACRIMA. Compared to out-of-domain pre-training or self-supervised pre-training, supervised pre-training achieved the best average performances on all of ResNet-18 (average AUROC = 0.857), DINOv2 (average AUROC = 0.788), and RETFound (average AUROC = 0.839). The best model is the ResNet-18 pre-trained with AIROGS diagnostic features, achieving the highest AUROC of 0.930 on the G1020 dataset, and an average AUROC of 0.857 across all datasets. These results underscore the superiority of incorporating domain-specific clinical labels to guide feature extraction, thereby improving model performance, and generalizability for glaucoma detection. Our findings advocate for the integration of supervised pre-training strategies into glaucoma detection model development, with significant potential to improve model performance and ultimately improve patient outcomes with decreased undiagnosed rate.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Autonomous screening for Diabetic Macular Edema using deep learning processing of retinal images 94%
- Quantification of Fundus Autofluorescence Features in a Molecularly Characterized Cohort of More Than 3500 Inherited Retinal Disease Patients from the United Kingdom 93%
- A Computational Framework for Intraoperative Pupil Analysis in Cataract Surgery 93%
Similar papers in this journal
- Equity-Enhanced Glaucoma Progression Prediction from OCT with Knowledge Distillation 97%
- A Deep Learning Based Smartphone Application for Early Detection of Nasopharyngeal Carcinoma Using Endoscopic Images 91%
- A human-in-the-loop explanation framework for morphologically transparent AI predictions from whole-slide images 90%
Similar papers in this journal
- Interpretable Detection of Epiretinal Membrane from Optical Coherence Tomography with Deep Neural Networks 96%
- Estimating Rates of Progression and Predicting Future Visual Fields in Glaucoma Using a Deep Variational Autoencoder 96%
- Circular functional analysis of OCT data for precise identification of structural phenotypes in the eye 95%
Similar papers in this journal
- An Inherently Interpretable AI model improves Screening Speed and Accuracy for Early Diabetic Retinopathy 97%
- Self-supervised contrastive learning improves machine learning discrimination of full thickness macular holes from epiretinal membranes in retinal OCT scans 97%
- Detecting papilloedema as a marker of raised intracranial pressure using artificial intelligence: a systematic review 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.