Generalizable AI predicts immunotherapy outcomes across cancers and treatments
SHEN, W.; Nguyen, T. H.; Li, M. M. R.; Huang, Y.; Moon, I.; Nair, N.; Marbach, D.; Zitnik, M.
Show abstract
Immune checkpoint inhibitors have become standard care across many cancers, but most patients do not respond. Predicting response remains challenging due to complex tumorimmune interactions and the poor generalizability of current biomarkers and models. Predictors such as tumor mutational burden, PD-L1 expression, and transcriptomic signatures often fail across cancer types, therapies, and clinical settings. There is a clear need for a robust, interpretable model that captures shared immune response principles and adapts to diverse clinical contexts. We present CO_SCPLOWOMPASSC_SCPLOW, a foundation model for predicting immunotherapy response from pan-cancer transcriptomic data using a concept bottleneck architecture. CO_SCPLOWOMPASSC_SCPLOWencodes tumor gene expression through 44 biologically grounded immune concepts representing immune cell states, tumor-microenvironment interactions, and signaling pathways. Trained on 10,184 tumors across 33 cancer types, CO_SCPLOWOMPASSC_SCPLOWoutperforms 22 baseline methods in 16 independent clinical cohorts spanning seven cancers and six immune checkpoint inhibitors, increasing precision by 8.5%, Matthews correlation coefficient by 12.3%, and area under the precision-recall curve by 15.7%, with minimal or no additional training. The model generalizes to unseen cancer types and treatments, supporting indication selection and patient stratification in early-phase clinical trials. Survival analysis shows that CO_SCPLOWOMPASSC_SCPLOW-stratified responders have significantly longer overall survival (hazard ratio = 4.7, p < 0.0001). Personalized response maps link gene expression to immune concepts, revealing distinct mechanisms of response and resistance. For example, among immune-inflamed non-responders, CO_SCPLOWOMPASSC_SCPLOWidentifies distinct resistance programs involving TGF-{beta} signaling, endothelial exclusion, CD4+ T cell dysfunction, and B cell deficiency. By combining mechanistic interpretability with transfer learning, CO_SCPLOWOMPASSC_SCPLOWprovides mechanistic insights into treatment response variability, supports clinical decision-making, and informs trial design.
Matching journals
The top 2 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- NeoPrecis: Enhancing Immunotherapy Response Prediction through Integration of Qualified Immunogenicity and Clonality-Aware Neoantigen Landscapes 98%
- Immune cell topography predicts response to PD-1 blockade in cutaneous T cell lymphoma 96%
- Community assessment of methods to deconvolve cellular composition from bulk gene expression 96%
Similar papers in this journal
- Swarm Learning as a privacy-preserving machine learning approach for disease classification 96%
- Functional interrogation uncovers a critical role for a high-plasticity cell state in lung adenocarcinoma 94%
- Large-scale clinical interpretation of genetic variants using evolutionary data and deep learning 94%
Similar papers in this journal
- Clonal transcriptomics identifies mechanisms of chemoresistance and empowers rational design of combination therapies. 94%
- Population Dynamics of Immunological Synapse Formation Induced by Bispecific T-cell Engagers Predict Clinical Pharmacodynamics and Treatment Resistance 94%
- Patient-derived xenografts and single-cell sequencing identifies three subtypes of tumor-reactive lymphocytes in uveal melanoma metastases 94%
Similar papers in this journal
Similar papers in this journal
- AI-Driven Predictive Biomarker Discovery with Contrastive Learning to Improve Clinical Trial Outcomes 97%
- Single-cell lineage and transcriptome reconstruction of metastatic cancer reveals selection of aggressive hybrid EMT states 95%
- Systematic Elucidation and Pharmacological Targeting of Tumor-Infiltrating Regulatory T Cell Master Regulators 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.