Impact of a machine learning-powered algorithm on pathologist HER2 IHC scoring in breast cancer
Shamshoian, J.; Shanis, Z.; Cabeen, R.; Yu, L.; Chakraborty, S.; Thibault, M.; Martin, B.; Padigela, H.; Juyal, D.; Javed, S. A.; Qian, W.; Kim, J.; Gerardin, Y.; Rucker, B.; Brosnan-Cashman, J.; Pokkalla, H.; Mehta, J.; Taylor-Weiner, A.; Walk, E.; Beck, A.; Montalto, M. C.; Glass, B.; Balasubramanian, S.
Show abstract
BackgroundHER2 expression level is a key factor in determining the optimal treatment course for breast cancer patients. Roughly 15% of breast cancers are HER2(+), and determination of HER2 status is routinely assessed by immunohistochemistry (IHC). Accurate assessment of the HER2 IHC score by pathologists is therefore critical, especially in light of novel therapeutic approaches demonstrating efficacy in the HER2-low setting. However, there is an opportunity to improve inter-pathologist agreement at the lower levels of HER2 scoring (0, 1+, and 2+). MethodsA machine learning model (AIM-HER2) was developed to generate accurate, slide-level HER2 scores aligned with ASCO-CAP guidelines in clinical breast cancer HER2 IHC specimens. AIM-HER2 was assessed as an AI-assist tool in a retrospective reader study, where 20 HER2-trained pathologists scored breast cancer cases (N=200) with and without AIM-HER2 assistance using a 2-cohort crossover design with a 3-week washout. A separate panel of 5 expert HER2 pathologists read all 200 cases manually to establish reference scores. ResultsIn a significant fraction of cases examined, less than 70% inter-pathologist agreement was observed. When used as an AI assist tool, AIM-HER2 improved inter-rater agreement overall and specifically at the 0/1+ and 1+/2+ cutoffs. Similarly, AIM-HER2 AI-assist significantly increased PPA at the 0/1+ and 1+/2+ cutoffs. When interacting with the AI-assist tool, pathologists displayed a wide range of override rates, and the quality of a pathologists overrides was correlated with their manual accuracy. Lastly, the impact of the reference panel on AIM-HER2 accuracy metrics was assessed, revealing that measurements of model accuracy are highly dependent on reference panel composition. ConclusionsThe use of AIM-HER2 as an AI-assist tool for scoring HER2 IHC in breast cancer may improve pathologist reproducibility and accuracy, particularly at the 0/1+ and 1+/2+ cutoffs.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Genomic Characterization of Lung Cancer in Never-Smokers Using Deep Learning 93%
- Tissue contamination challenges the credibility of machine learning models in real world digital pathology 92%
- Attention-based whole-slide image compression achieves pathologist-level pre-screening of multi-organ routine histopathology biopsies 92%
Similar papers in this journal
- Quantification of HER2-low and ultra-low expression in breast cancer specimens by quantitative IHC and artificial intelligence 94%
- Independent assessment of a deep learning system for lymph node metastasis detection on the Augmented Reality Microscope 93%
- Development of an Interactive Web Dashboard to Facilitate the Reexamination of Pathology Reports for Instances of Underbilling of CPT Codes 92%
Similar papers in this journal
- Computer Vision Identifies Recurrent and Non-Recurrent Ductal Carcinoma in situ Lesions with Special Emphasis on African American Women 94%
- Non-Metastatic Axillary Lymph Nodes Have Distinct Morphology and Immunophenotype in Obese Breast Cancer patients at Risk for Metastasis 93%
- Detection of Colorectal Adenocarcinoma and Grading Dysplasia on Histopathologic Slides Using Deep Learning 92%
Similar papers in this journal
- RNA Sequencing-Based Single Sample Predictors of Molecular Subtype and Risk of Recurrence for Clinical Assessment of Early-Stage Breast Cancer 94%
- Unmasking the tissue microecology of ductal carcinoma in situ with deep learning 93%
- Predicting neoadjuvant chemotherapy benefit using deep learning from stromal histology in breast cancer 91%
Similar papers in this journal
- Clinical validation of Whole Genome Sequencing for routine cancer diagnostics 92%
- PanelCAT: an Open-Source Comparative Analysis Tool for Next-Generation Sequencing Panel Target Regions 91%
- Oncogenicity Variant Interpreter (OncoVI): oncogenicity guidelines implementation to support somatic variants interpretation in precision oncology 91%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.