High content image analysis in routine diagnostic histopathology predicts outcomes in HPV-associated oropharyngeal squamous cell carcinomas
Hue, J.; Valinciute, Z.; Thavaraj, S.; Veschini, L.
Show abstract
ObjectiveRoutine haematoxylin and eosin (H&E) photomicrographs from human papillomavirus-associated oropharyngeal squamous cell carcinomas (HPV+OpSCC) contain a wealth of prognostic information. In this study, we set out to develop a high content image analysis workflow to quantify features of H&E images from HPV+OpSCC patients to identify prognostic features which can be used for prediction of patient outcomes. MethodsWe have developed a dedicated image analysis workflow using open-source software, for single-cell segmentation and classification. This workflow was applied to a set of 567 images from diagnostic H&E slides in a retrospective cohort of HPV+OpSCC patients with favourable (n = 29) and unfavourable (n = 29) outcomes. Using our method, we have identified 31 quantitative prognostic features which were quantified in each sample and used to train a neural network model to predict patient outcomes. The model was validated by k-fold cross-validation using 10 folds and a test set. ResultsUnivariate and multivariate statistical analyses revealed significant differences between the two patient outcome groups in 31 and 16 variables respectively (P<0.05). The neural network model had an overall accuracy of 78.8% and 77.7% in recognising favourable and unfavourable prognosis patients when applied to the test set and k-fold cross-validation respectively. ConclusionOur open-source H&E analysis workflow and model can predict HPV+OpSCC outcomes with promising accuracy. Our work supports the use of machine learning in digital pathology to exploit clinically relevant features in routine diagnostic pathology without additional biomarkers.
Matching journals
The top 8 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Artificial Intelligence for Advance Requesting of Immunohistochemistry in Diagnostically Uncertain Prostate Biopsies 92%
- Attention-based whole-slide image compression achieves pathologist-level pre-screening of multi-organ routine histopathology biopsies 91%
- MIXTURE of human expertise and deep learning—Developing an explainable model for predicting pathological diagnosis and survival in patients with interstitial lung disease 91%
Similar papers in this journal
- Automated Clear Cell Renal Carcinoma Grade Classification with Prognostic Significance 94%
- Pixelwise H-score: a novel digital image analysis based-metric to quantify membrane biomarker expression from immunohistochemistry images 93%
- Weakly supervised learning for multi-organ adenocarcinoma classification in whole slide images 92%
Similar papers in this journal
- Inference of core needle biopsy whole slide images requiring definitive therapy for prostate cancer 92%
- Deep learning-based tumor microenvironment segmentation is predictive of tumor mutations and patient survival in non-small-cell lung cancer 90%
- Simplified Molecular Classification of Lung Adenocarcinomas Based on EGFR, KRAS, and TP53 Mutations 90%
Similar papers in this journal
- Identifying relationships between imaging phenotypes and lung cancer-related mutation status: EGFR and KRAS 92%
- PathProfiler: Automated Quality Assessment of Retrospective Histopathology Whole-Slide Image Cohorts by Artificial Intelligence, A Case Study for Prostate Cancer Research 92%
- Automated and Manual Quantification of Tumour Cellularity in Digital Slides for Tumour Burden Assessment 92%
Similar papers in this journal
- A digital score of peri-epithelial lymphocytic activity predicts malignant transformation in oral epithelial dysplasia 97%
- Spatial Effects of Infiltrating T cells on Neighbouring Cancer Cells and Prognosis in Stage III CRC patients 92%
- Development of a multi-scanner facility for data acquisition for digital pathology artificial intelligence 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.