Counterfactual Diffusion Models for Mechanistic Explainability of Artificial Intelligence Models in Pathology
Zigutyte, L.; Lenz, T.; Han, T.; Hewitt, K. J.; Reitsam, N. G.; Foersch, S.; Carrero, Z.; Unger, M.; Pearson, A. T.; Truhn, D.; Kather, J. N.
Show abstract
Deep learning can extract predictive and prognostic biomarkers from histopathology whole slide images. However, explainable artificial intelligence approaches widely used in digital pathology, such as attention heatmaps and class activation mapping, offer only limited interpretability regarding the features captured by classifiers. Here, we present MoPaDi (Morphing histoPathology Diffusion), a framework for generating counterfactual explanations for histopathology images that reveal which morphological or style features drive classifier predictions. MoPaDi combines diffusion autoencoders with task-specific multiple instance learning classifiers to manipulate images and flip predictions by modifying relevant features. We evaluated the framework on multiple datasets spanning colorectal, breast, liver, and lung cancers, including tissue type, cancer subtype, and biomarker (microsatellite instability) classification tasks. We assessed counterfactual explanations through quantitative analyses, pathologists evaluations, and independent foundation model-based classifiers. We found that MoPaDi was able to generate realistic counterfactual histopathology images, enabling pathologists to identify morphological features associated with the change in model predictions. Unlike conventional reviews of highly attended regions typical in digital pathology, MoPaDi explanations enabled pathologists to directly identify morphological features driving the classifiers predictions from a limited number of top-contributing tiles. Consistent with the literature, our biomarker classifier associated high microsatellite instability with mucinous differentiation, glandular patterns, and lymphocytic infiltration. Furthermore, MoPaDi revealed that changes in classifier predictions were mainly driven by morphological alterations rather than staining differences. Overall, MoPaDi is a practical framework for counterfactual explanations in computational pathology that reveals model-specific drivers of classification and increases trust in deep learning models.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- STAIG: Spatial Transcriptomics Analysis via Image-Aided Graph Contrastive Learning for Domain Exploration and Alignment-Free Integration 97%
- PHARAOH: A collaborative crowdsourcing platform for PHenotyping And Regional Analysis Of Histology 96%
- Compressed sensing expands the multiplexity of imaging mass cytometry 95%
Similar papers in this journal
- Identifying perturbations that boost T-cell infiltration into tumours via counterfactual learning of their spatial proteomic profiles 96%
- Ultra-fast Prediction of Somatic Structural Variations by Reduced Read Mapping via Pan-Genome k-mer Sets 93%
- Single-shot 3D photoacoustic tomography using a single-element detector for ultrafast imaging of hemodynamics 93%
Similar papers in this journal
- Predicting MammaPrint Recurrence Risk from Breast Cancer Pathological Images Using a Weakly Supervised Transformer 97%
- Information-Distilled Generative Label-Free Morphological Profiling Encodes Cellular Heterogeneity 95%
- Unsupervised segmentation of 3D microvascular photoacoustic images using deep generative learning 95%
Similar papers in this journal
- Generative Adversarial Networks Accurately Reconstruct Pan-Cancer Histology from Pathologic, Genomic, and Radiographic Latent Features 97%
- Biologically relevant integration of transcriptomics profiles from cancer cell lines, patient-derived xenografts and clinical tumors using deep learning 96%
- Charting the transcriptomic landscape of primary and metastatic cancers in relation to their origin and target normal tissues 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.