OphthUS-GPT: Multimodal AI for Automated Reporting in Ophthalmic B-Scan Ultrasound
Gan, F.; Chen, l.; Qin, W. g.; Han, Q. l.; Long, X.; Fan, H. m.; li, X. y.; Yu, H. z.; Zhang, J. h.; Xu, N.; Cheng, J. x.; Cao, J.; Liu, K. c.; Shao, Y. n.; Li, X. n.; Wan, Q.; Liu, T.; You, Z. p.
Show abstract
IMPORTANCEThe rapid advancement of AI in ophthalmology is transforming diagnostics, especially in resource-limited settings. The shortage of ophthalmologists and lack of standardized reporting creates an urgent need for AI systems capable of automated reporting and interactive decision support. OBJECTIVETo develop OphthUS-GPT, a multimodal AI system integrating BLIP and DeepSeek models for automated report generation and clinical decision support from ophthalmic B-scan ultrasound images. DESIGN, SETTING, AND PARTICIPANTSThis retrospective study at the Affiliated Eye Hospital of Jiangxi Medical College collected B-scan ultrasound reports between 2017-2024, including 54,696 images and 9,392 reports from 31,943 patients (mean age 49.14{+/-}0.124 years, 50.15% male). MAIN OUTCOMES AND MEASURESEvaluation included two components: diagnostic report generation and question-answering system assessment. Report generation was evaluated using text metrics (ROUGE-L, CIDEr), disease classification metrics (accuracy, sensitivity, specificity, precision, F1 score), and ophthalmologist ratings for accuracy and completeness. The question-answering system was assessed by ophthalmologists rating answers on accuracy, completeness, potential harm, and satisfaction. RESULTSOphthUS-GPT achieved ROUGE-L and CIDEr scores of 0.6131 and 0.9818 in report generation. For common conditions, accuracy exceeded 90% with precision >70%. Expert assessment showed >90% of reports scored [≥] 3/5 for correctness and 96% for completeness. The DeepSeek-R1-Distill-Llama-8B (DeepSeek) question-answering component performed comparably to GPT4o and OpenAI-o1, outperforming other models. CONCLUSIONS AND RELEVANCOphthUS-GPT demonstrated excellent performance in automatic report generation and intelligent Q&A, offering a novel solution for ophthalmic ultrasound interpretation and clinical decision support.
Matching journals
The top 2 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Equity-Enhanced Glaucoma Progression Prediction from OCT with Knowledge Distillation 95%
- A Deep Learning Based Smartphone Application for Early Detection of Nasopharyngeal Carcinoma Using Endoscopic Images 93%
- The clinician-AI interface: intended use and explainability in FDA-cleared AI devices for medical image interpretation 93%
Similar papers in this journal
Similar papers in this journal
- An Inherently Interpretable AI model improves Screening Speed and Accuracy for Early Diabetic Retinopathy 95%
- Self-supervised contrastive learning improves machine learning discrimination of full thickness macular holes from epiretinal membranes in retinal OCT scans 93%
- Development and Validation of a Deep Learning Model for Detecting Signs of Tuberculosis on Chest Radiographs among US-bound Immigrants and Refugees 92%
Similar papers in this journal
- AutoMorph: Automated Retinal Vascular Morphology Quantification via a Deep Learning Pipeline 95%
- Advancing Question-Answering in Ophthalmology with Retrieval Augmented Generations (RAG): Benchmarking Open-source and Proprietary Large Language Models 93%
- Current applications of artificial intelligence for Fuchs endothelial corneal dystrophy: a systematic review 92%
Similar papers in this journal
- Performance of DeepSeek-R1 in Ophthalmology: An Evaluation of Clinical Decision-Making and Cost-Effectiveness 96%
- Unveiling the Clinical Incapabilities: A Benchmarking Study of GPT-4V(ision) for Ophthalmic Multimodal Image Analysis 95%
- Autonomous Screening for Laser Photocoagulation in Fundus Images Using Deep Learning 91%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.