A novel workflow for the safe and effective integration of AI as supporting reader in double reading breast cancer screening: A large-scale retrospective evaluation
Ng, A. Y.; Glocker, B.; Oberije, C.; Fox, G.; Nash, J.; Karpati, E.; Kerruish, S.; Kecskemethy, P. D.
Show abstract
ObjectivesTo evaluate the effectiveness of a novel strategy for using AI as a supporting reader for the detection of breast cancer in mammography-based double reading screening practice. Instead of replacing a human reader, here AI serves as the second reader only if it agrees with the recall/no-recall decision of the first human reader. Otherwise, a second human reader makes an assessment, enacting standard human double reading. DesignRetrospective large-scale, multi-site, multi-device, evaluation study. Participants280,594 cases from 180,542 female participants who were screened for breast cancer with digital mammography between 2009 and 2019 at seven screening sites in two countries (UK and Hungary). Main outcome measuresPrimary outcome measures were cancer detection rate, recall rate, sensitivity, specificity, and positive predictive value. Secondary outcome was reduction in workload measured as arbitration rate and number of cases requiring second human reading. ResultsThe novel workflow was found to be superior or non-inferior on all screening metrics, almost halving arbitration and reducing the number of cases requiring second human reading by up to 87.50% compared to human double reading. ConclusionsAI as a supporting reader adds a safety net in case of AI discordance compared to alternative workflows where AI replaces the second human reader. In the simulation using large-scale historical data, the proposed workflow retains screening performance of the standard of care of human double reading while drastically reducing the workload. Further research should study the impact of the change in case mix for the second human reader as they would only assess cases where the AI and first human reader disagree.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Classification performance bias between training and test sets in a limited mammography dataset 94%
- Early user experience and lessons learned using ultra-portable digital X-ray with computer-aided detection (DXR-CAD) products: A qualitative study from the perspective of healthcare providers 92%
- Navigated ultrasound bronchoscopy with integrated positron emission tomography - A human feasibility study 92%
Similar papers in this journal
- A Machine Learning Ensemble Based on Radiomics to Predict BI-RADS Category and Reduce the Biopsy Rate of Ultrasound-Detected Suspicious Breast Masses 94%
- The NILS study protocol - a retrospective validation study of a preoperative decision-making tool for non-invasive lymph node staging in women with primary breast cancer [ISRCTN14341750] 92%
- Auto-detection of motion artifacts on CT pulmonary angiograms with a physician-trained AI algorithm 92%
Similar papers in this journal
- Automated and Manual Quantification of Tumour Cellularity in Digital Slides for Tumour Burden Assessment 94%
- Reproducible And Clinically Translatable Deep Neural Networks For Cervical Screening 93%
- Content-based image retrieval assists radiologists in diagnosing eye and orbital mass lesions in MRI 93%
Similar papers in this journal
- Model uncertainty estimates for deep learning mammographic density prediction using ordinal and classification approaches 94%
- Breast density prediction from low and standard dose mammograms using deep learning: effect of image resolution and model training approach on prediction quality 91%
- Mammographic density assessed using deep learning in women at high risk of developing breast cancer: the effect of weight change on density 91%
Similar papers in this journal
- Development and validation of multivariable machine learning algorithms to predict risk of cancer in symptomatic patients referred urgently from primary care 94%
- Retrospective Validation of an Artificial Intelligence System for Diagnostic Assessment of Prostate Biopsies on the ProMort Cohort: Study Protocol 91%
- Large language model-based information extraction from free-text radiology reports: a scoping review protocol 90%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.