Back

Flexible and Highly-Efficient Feature Perception for Molecular Traits Prediction via Self-interactive Deep Learning

Hu, Y.; Sirinukunwattana, K.; Li, B.; Gaitskell, K.; Bonnaffe, W.; Wojciechowska, M.; Wood, R.; Alham, N. K.; Malacrino, S.; Woodcock, D.; Verrill, C.; Ahmed, A.; Rittscher, J.

2023-08-05 pathology
10.1101/2023.07.30.23293391 medRxiv
Show abstract

Predicting disease-related molecular traits from histomorphology brings great opportunities for precision medicine. Despite the rich information present in histopathological images, extracting fine-grained molecular features from standard whole slide images (WSI) is non-trivial. The task is further complicated by the lack of annotations for subtyping and contextual histomorphological features that might span multiple scales. This work proposes a novel multiple-instance learning (MIL) framework capable of WSI-based cancer morpho-molecular subtyping across scales. Our method, debuting as Inter-MIL, follows a weakly-supervised scheme. It enables the training of the patch-level encoder for WSI in a task-aware optimisation procedure, a step normally improbable in most existing MIL-based WSI analysis frameworks. We demonstrate that optimising the patch-level encoder is crucial to achieving high-quality fine-grained and tissue-level subtyping results and offers a significant improvement over task-agnostic encoders. Our approach deploys a pseudo-label propagation strategy to update the patch encoder iteratively, allowing discriminative subtype features to be learned. This mechanism also empowers extracting fine-grained attention within image tiles (the small patches), a task largely ignored in most existing weakly supervised-based frameworks. With Inter-MIL, we carried out four challenging cancer molecular subtyping tasks in the context of ovarian, colorectal, lung, and breast cancer. Extensive evaluation results show that Inter-MIL is a robust framework for cancer morpho-molecular subtyping with superior performance compared to several recently proposed methods, even in data-limited scenarios where the number of available training slides is less than 100. The iterative optimisation mechanism of Inter-MIL significantly improves the quality of the image features learned by the patch embedded and generally directs the attention map to areas that better align with experts interpretation, leading to the identification of more reliable histopathology biomarkers.

Matching journals

The top 3 journals account for 50% of the predicted probability mass.

1
Nature Machine Intelligence
70 papers in training set
Top 0.1%
30.7%
2
Nature Communications
5641 papers in training set
Top 9%
18.3%
3
Medical Image Analysis
35 papers in training set
Top 0.1%
7.2%
50% of probability mass above
4
Advanced Intelligent Systems
11 papers in training set
Top 0.1%
6.6%
5
npj Digital Medicine
118 papers in training set
Top 1%
3.4%
6
iScience
1154 papers in training set
Top 7%
3.2%
7
Modern Pathology
22 papers in training set
Top 0.2%
2.1%
8
Nature Methods
385 papers in training set
Top 4%
2.1%
9
Scientific Reports
3612 papers in training set
Top 48%
2.1%
10
Advanced Science
286 papers in training set
Top 5%
1.7%
11
Briefings in Bioinformatics
354 papers in training set
Top 5%
1.7%
12
Communications Biology
993 papers in training set
Top 17%
1.5%
13
npj Precision Oncology
53 papers in training set
Top 1.0%
1.5%
14
Cell Reports Medicine
153 papers in training set
Top 3%
1.4%
15
PLOS Computational Biology
1863 papers in training set
Top 16%
1.4%
16
eBioMedicine
183 papers in training set
Top 4%
1.1%
17
Communications Medicine
113 papers in training set
Top 4%
1.1%
18
Nature Medicine
125 papers in training set
Top 3%
1.0%
19
IEEE Journal of Biomedical and Health Informatics
37 papers in training set
Top 1%
0.8%
20
Genome Biology
637 papers in training set
Top 10%
0.6%
21
Journal of Pathology Informatics
15 papers in training set
Top 0.3%
0.6%
22
Light: Science & Applications
16 papers in training set
Top 0.4%
0.6%
23
Bioinformatics
1204 papers in training set
Top 9%
0.6%
24
Bioinformatics Advances
203 papers in training set
Top 5%
0.6%