Back

Deep learning for rapid and reproducible histology scoring of lung injury in a porcine model

Neves Da Silva, I. A.; Kazemi Rashed, S.; Hedlund, L.; Lidfeldt, A.; Gvazava, N.; Stegmayr, J.; Skoryk, V.; Aits, S.; Wagner, D. E.

2023-05-14 bioinformatics
10.1101/2023.05.12.540340 bioRxiv
Show abstract

Acute respiratory distress syndrome (ARDS) is a life-threatening condition with mortality rates between 30-50%. Although in vitro models replicate some aspects of ARDS, small and large animal models remain the primary research tools due to the multifactorial nature of the disease. When using these animal models, histology serves as the gold standard method to confirm lung injury and exclude other diagnoses as high-resolution chest images are often not feasible. Semi-quantitative scoring performed by independent observers is the most common form of histologic analysis in pre-clinical animal models of ARDS. Despite progress in standardizing analysis procedures, objectively comparing histological injuries remains challenging, even for highly-trained pathologists. Standardized scoring simplifies the task and allows better comparisons between research groups and across different injury models, but it is time-consuming, and interobserver variability remains a significant concern. Convolutional neural networks (CNNs), which have emerged as a key tool in image analysis, could automate this process, potentially enabling faster and more reproducible analysis. Here we explored the reproducibility of human standardized scoring for an animal model of ARDS and its suitability for training CNNs for automated scoring at the whole slide level. We found large variations between human scorers, even for pre-clinical experts and board-certified pathologies in evaluating ARDS animal models. We demonstrate that CNNs (VGG16, EfficientNetB4) are suitable for automated scoring and achieve up to 83% F1-score and 78% accuracy. Thus, CNNs for histopathological classification of acute lung injury could help reduce human variability and eliminate a time-consuming manual research task with acceptable performance.

Matching journals

The top 10 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.