Back

Deep Learning in Automating Breast Cancer Diagnosis from Microscopy Images

Gu, Q.; Prodduturi, N.; Hart, S. N.

2023-06-16 pathology
10.1101/2023.06.15.23291437 medRxiv
Show abstract

ContextBreast cancer is one of the most common cancers in women. With early diagnosis, some breast cancers are highly curable. However, the concordance rate of breast cancer diagnosis from histology slides by pathologists is unacceptably low. Classifying normal versus tumor breast tissues from microscopy images of breast histology is an ideal case to use for deep learning and could help to more reproducibly diagnose breast cancer. Since data preprocessing and hyperparameter configurations have impacts on breast cancer classification accuracies of deep learning models, training a deep learning classifier with appropriate data preprocessing approaches and optimized hyperparameter configurations could improve breast cancer classification accuracy. Methods and MaterialUsing 12 combinations of deep learning model architectures (i.e., including 5 non-specialized and 7 digital pathology-specialized model architectures), image data preprocessing, and hyperparameter configurations, the validation accuracy of tumor versus normal classification were calculated using the BreAst Cancer Histology (BACH) dataset. ResultsThe DenseNet201, a non-specialized model architecture, with transfer learning approach achieved 98.61% validation accuracy compared to only 64.00% for the digital pathology-specialized model architecture. ConclusionsThe combination of image data preprocessing approaches and hyperparameter configurations have a profound impact on the performance of deep neural networks for image classification. To identify a well-performing deep neural network to classify tumor versus normal breast histology, researchers should not only focus on developing new models specifically for digital pathology, since hyperparameter tuning for existing deep neural networks in the computer vision field could also achieve a high (often better) prediction accuracy.

Matching journals

The top 5 journals account for 50% of the predicted probability mass.

1
Journal of Pathology Informatics
15 papers in training set
Top 0.1%
32.9%
2
Scientific Reports
3612 papers in training set
Top 9%
7.2%
3
Diagnostics
50 papers in training set
Top 0.4%
4.8%
4
The American Journal of Pathology
32 papers in training set
Top 0.1%
4.3%
5
PLOS ONE
5266 papers in training set
Top 38%
3.2%
50% of probability mass above
6
Laboratory Investigation
13 papers in training set
Top 0.1%
3.2%
7
Cancers
213 papers in training set
Top 2%
3.1%
8
Advanced Intelligent Systems
11 papers in training set
Top 0.1%
3.1%
9
Modern Pathology
22 papers in training set
Top 0.1%
3.1%
10
Journal of Medical Imaging
11 papers in training set
Top 0.1%
3.1%
11
Nature Machine Intelligence
70 papers in training set
Top 1%
1.7%
12
GigaScience
212 papers in training set
Top 3%
1.7%
13
Nature Communications
5641 papers in training set
Top 47%
1.5%
14
JNCI Cancer Spectrum
10 papers in training set
Top 0.2%
1.5%
15
Scientific Data
209 papers in training set
Top 2%
1.3%
16
Journal of the American Society for Mass Spectrometry
37 papers in training set
Top 0.4%
1.1%
17
Biology Methods and Protocols
61 papers in training set
Top 1%
1.1%
18
The Lancet Digital Health
25 papers in training set
Top 0.5%
1.1%
19
Breast Cancer Research
36 papers in training set
Top 0.4%
1.1%
20
npj Precision Oncology
53 papers in training set
Top 1%
1.0%
21
Neuropathology and Applied Neurobiology
15 papers in training set
Top 0.4%
0.8%
22
Cells
249 papers in training set
Top 7%
0.8%
23
eBioMedicine
183 papers in training set
Top 6%
0.8%
24
Cancer Research Communications
51 papers in training set
Top 2%
0.8%
25
npj Digital Medicine
118 papers in training set
Top 3%
0.8%
26
BMC Cancer
67 papers in training set
Top 2%
0.6%
27
Genomics, Proteomics & Bioinformatics
172 papers in training set
Top 2%
0.6%
28
Communications Biology
993 papers in training set
Top 35%
0.6%
29
Light: Science & Applications
16 papers in training set
Top 0.3%
0.6%
30
Bioinformatics
1204 papers in training set
Top 9%
0.6%