Back

DeepPathway: Predicting Pathway Expression from Histopathology Images

Ahsan, M. A.; Piper Hanley, K.; Fergie, M.; O'leary, C.; Borst, G.; Roncaroli, F.; Rattray, M.; Iqbal, M.; Baker, S. M.

2025-07-25 bioinformatics
10.1101/2025.07.21.665956 bioRxiv
Show abstract

Spatial transcriptomics (ST) technologies provide spatially resolved gene expression along with image data, allowing the integrative analysis of complex tissue microenvironments. Despite their potential, the widespread adoption of ST remains limited due to high costs, and methodological challenges in data acquisition. Thus, there have been recent efforts to develop deep learning methods capable of inferring spatial gene expression from the much cheaper and easily available haematoxylin and eosin (H&E) images. These methods demonstrate promising results in reconstructing transcriptomic landscapes within tissue sections. While existing approaches predominantly focus on gene-level predictions, biological processes are often regulated at the pathway level through coordinated activity among functionally related genes. We present DeepPathway, a contrastive learning-based approach trained on ST data to predict pathway expression from H&E-stained sections. We compute input pathway expression by summarizing the expression of constituent genes using established pathway definitions. We evaluate the performance of our method on two prostate cancer datasets and validate our approach on the H&E images acquired from The Cancer Genome Atlas (TCGA) clearly differentiating between normal and tumour tissues. Finally, we apply our method to predict hypoxia signatures using H&Es of brain tumour samples where hypoxia staining with pimonidazole was available as ground truth. Implementation code for DeepPathway is available at https://github.com/aahsan045/DeepPathway.

Matching journals

The top 5 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.