Back

Identification of High-Risk Cells in Single-Cell Spatially Resolved Transcriptomics Data Using Deep Transfer Learning

Chatterjee, D.; Couetil, J. L.; Liu, Z.; Dong, T.; Huang, K.; Chen, C.; Zhang, J.; Kalwat, M. A.; Johnson, T. S.

2025-02-05 bioinformatics
10.1101/2025.01.30.635803 bioRxiv
Show abstract

SummaryThe examination of high-risk cells and regions in tissue samples from spatially resolved transcriptomics platforms offers meaningful insights into specific disease processes. For existing methods, while cell types or clusters can be identified and associated with disease attributes, individual cells are unable to be associated in the same manner. MethodDiag-nostic Evidence GAuge of Single-cells and spatial transcriptomics (DEGAS), solves the above problem by employing latent representations of gene expression data and domain adaptation to transfer disease attributes from patients to individual cells from single-cell RNA sequencing datasets. In this research, we present and evaluate DEGASs versatility in adapting to data arising from various single-cell spatially resolved transcriptomics (scSRT) platforms. DEGAS successfully identified high-risk cells and regions in liver hepatocellular carcinoma and skin cutaneous melanoma, which were validated through known markers. Additionally, DEGAS was applied to our newly generated Type II Diabetes Xenium dataset revealing high-risk cells within the tissue samples. Availability and ImplementationThe DEGAS software can be accessed at https://github.com/tsteelejohnson91/DEGAS. For the updated smoothing functions and associated codes, visit https://github.com/dchatter04/DEGAS-Spatial-Smoothing. Sources for the datasets reviewed are detailed in their respective sections. A description of some datasets, along with extra tables and figures, is provided in the Supplementary Materials file. Our newly generated Xenium data for Type II Diabetes can be found at https://doi.org/10.7303/syn68699752.

Published in Bioinformatics (predicted rank #1) · training set

Matching journals

The top 4 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.