AI-driven Discovery of Morphomolecular Signatures in Toxicology
Jaume, G.; Peeters, T.; Song, A. H.; Pettit, R.; Williamson, D. F. K.; Oldenburg, L.; Vaidya, A.; De Brot, S.; Chen, R. J.; Thiran, J.-P.; Le, L. P.; Gerber, G.; Mahmood, F.
Show abstract
Early identification of drug toxicity is essential yet challenging in drug development. At the preclinical stage, toxicity is assessed with histopathological examination of tissue sections from animal models to detect morphological lesions. To complement this analysis, toxicogenomics is increasingly employed to understand the mechanism of action of the compound and ultimately identify lesion-specific safety biomarkers for which in vitro assays can be designed. However, existing works that aim to identify morphological correlates of expression changes rely on qualitative or semi-quantitative morphological characterization and remain limited in scale or morphological diversity. Artificial intelligence (AI) offers a promising approach for quantitatively modeling this relationship at an unprecedented scale. Here, we introduce GEESE, an AI model designed to impute morphomolecular signatures in toxicology data. Our model was trained to predict 1,536 gene targets on a cohort of 8,231 hematoxylin and eosin-stained liver sections from Rattus norvegicus across 127 preclinical toxicity studies. The model, evaluated on 2,002 tissue sections from 29 held-out studies, can yield pseudo-spatially resolved gene expression maps, which we correlate with six key drug-induced liver injuries (DILI). From the resulting 25 million lesion-expression pairs, we established quantitative relations between up and downregulated genes and lesions. Validation of these signatures against toxicogenomic databases, pathway enrichment analyses, and human hepatocyte cell lines asserted their relevance. Overall, our study introduces new methods for characterizing toxicity at an unprecedented scale and granularity, paving the way for AI-driven discovery of toxicity biomarkers. Live demo: https://mahmoodlab.github.io/tox-discovery-ui/
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
- ProT-Diff: A Modularized and Efficient Approach to De Novo Generation of Antimicrobial Peptide Sequences through Integration of Protein Language Model and Diffusion Model 94%
- Chronic opioid treatment arrests neurodevelopment and alters synaptic activity in human midbrain organoids 93%
- Predicting MammaPrint Recurrence Risk from Breast Cancer Pathological Images Using a Weakly Supervised Transformer 93%
Similar papers in this journal
- Large Scale Cell Painting Guided Compound Selection Reveals Activity Cliffs and Functional Relationships 95%
- The pharmacoepigenomic landscape of cancer cell lines reveals the epigenetic component of drug sensitivity 93%
- Murine breast cancers disorganize the liver transcriptome in zonated manners 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.