Star-Motifs: Revealing Single-Cell Spatiotypes from Routine Histology Using Star-Convex Neighborhoods
Gunesli, G. N.; Dawood, M.; Young, L. S.; Minhas, F.; Raza, S. E. A.; Rajpoot, N. M.
Show abstract
The tumor microenvironment (TME) is a dynamic interplay among cancer, immune, and stromal cells that profoundly influences tumor growth, progression, and treatment response. Although recent spatial transcriptomics and multiplex imaging technologies offer fine-grained insights into the TME, their high cost and lengthy turnaround times limit routine clinical use. In contrast, whole slide images (WSIs) of Hematoxylin and eosin (H&E)-stained are readily available for most patients and provide valuable information about tumor biology. Here, we present a method that explicitly models each cells local neighborhood to identify recurrent spatial patterns, which we call "Star-Motifs." We first encode the local microenvironment of each cell using a "Star-Environment" representation that captures distances and angles to multiple surrounding cell types. We then perform unsupervised clustering of these descriptors to discover distinct Star-Motifs. We validated our method on 1.8 billion cells across 3,105 diagnostic slides from The Cancer Genome Atlas, spanning five distinct cell types: neoplastic, inflammatory, stromal, necrotic, and non-neoplastic. Our method uncovered a diverse range of Star-Motifs, from densely packed tumor clusters to tumor cells surrounded by immune cells (tumor-infiltrating lymphocytes). At the patient level, the relative abundances of these Star-Motifs correlate with patient survival outcomes, immune subtypes, and molecular alterations, offering interpretable, clinically relevant insights. Moreover, classical machine learning models trained on these abundances match or surpass deep learning approaches in predicting overall survival and immune-related biomarkers while remaining interpretable. By capturing higher-order spatial arrangements from routinely available histology slides, Star-Motifs provides a scalable, transparent platform for TME profiling, biomarker discovery, and personalized risk assessment in oncology.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- MorphLink: Bridging Cell Morphological Behaviors and Molecular Dynamics in Multi-modal Spatial Omics 97%
- METI: Deep profiling of tumor ecosystems by integrating cell morphology and spatial transcriptomics 97%
- Self-Supervised Learning Reveals Clinically Relevant Histomorphological Patterns for Therapeutic Strategies in Colon Cancer 96%
Similar papers in this journal
- Integrative pan-cancer analysis reveals a common architecture of dysregulated transcriptional networks characterized by loss of enhancer methylation 95%
- Inferring latent temporal progression and regulatory networks from cross-sectional transcriptomic data of cancer samples 95%
- Randomized Spatial PCA (RASP): a computationally efficient method for dimensionality reduction of high-resolution spatial transcriptomics data 94%
Similar papers in this journal
- Image-Based Consensus Molecular Subtyping in Rectal Cancer Biopsies and Response to Neoadjuvant Chemoradiotherapy 95%
- Generalizing AI-driven Assessment of Immunohistochemistry across Immunostains and Cancer Types: A Universal Immunohistochemistry Analyzer 95%
- Predicting the Tumor Microenvironment Composition and Immunotherapy Response in Non-Small Cell Lung Cancer from Digital Histopathology Images 94%
Similar papers in this journal
Similar papers in this journal
- Diagnostic Evidence GAuge of Single cells (DEGAS): A flexible deep-transfer learning framework for prioritizing cells in relation to disease 95%
- DeepProg: an ensemble of deep-learning and machine-learning models for prognosis prediction using multi-omics data 95%
- Pan-cancer identification of clinically relevant genomic subtypes using outcome-weighted integrative clustering 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.