A Multi-Omics Computational Pipeline for Systematic Discovery of Retired Self-Antigens as Cancer Vaccine Targets
Wang, V.; Deng, S.; Aguilar, R.
Show abstract
BackgroundThe retired antigen hypothesis, introduced by Tuohy and colleagues, proposes that tissue-specific proteins expressed conditionally during early life or reproductive stages, then silenced in normal aging tissue, represent safe and effective cancer vaccine targets when re-expressed in tumors. To date, discovery of retired antigens has relied entirely on hypothesis-driven wet lab work, limiting throughput. MethodsHere we present RADAR (Retired Antigen Discovery and Ranking), a multi-omics computational pipeline implemented on a standard server that systematically identifies retired antigen candidates. RADAR comprises four core discovery layers integrating: 1) The Genotype-Tissue Expression Portal (GTEx) normal tissue expression, 2) TCGA tumor re-expression, 3) DNA methylation, and 4) miRNA regulatory networks, each applied sequentially to identify genes exhibiting the epigenetic and post-transcriptional hallmarks of tissue-specific retirement followed by tumor re-activation. Candidate characterization is further supported by three automated modules: 1) protein-level safety screening via the Human Protein Atlas, 2) molecular subtype enrichment analysis, and 3) cross-cancer confirmation, which execute automatically when the relevant data are available for the selected cancer type. ResultsThe pipeline independently validated known targets including alpha-lactalbumin (LALBA, the basis of the Tuohy Phase 1 triple-negative breast cancer vaccine trial) and anti-Mullerian hormone (AMH), consistent with Tuohys ovarian cancer vaccine program targeting AMHR2, and rediscovered multiple known cancer-testis antigens (MAGEA1, MAGEC1, SSX1) as positive controls. Among 4,664 initial candidates derived from GTEx, the pipeline identified 20 high-confidence retired antigen candidates passing all filters. DCAF4L2, COX7B2, TEX19, and CT83 emerge as the highest-priority novel candidates for experimental validation, demonstrating zero expression in critical somatic organs, strong epigenetic silencing, and significant re-expression across multiple cancer types. ConclusionRADAR provides the first systematic computational framework for retired antigen discovery, offering a reproducible and scalable approach to expanding the cancer immunoprevention pipeline beyond individually characterized targets. The pipeline is fully reproducible, requires no specialized hardware, and is immediately extensible to additional TCGA cancer types.
Matching journals
The top 10 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Genome profiles of lymphovascular breast cancer cells reveal multiple clonally differentiated outcomes with multi-regional LCM and G&T-seq 92%
- Evaluation of homologous recombination repair status in metastatic prostate cancer by next-generation sequencing and functional tissue-based immunofluorescence assays 92%
- Epitopes targeted by T cells in convalescent COVID-19 patients 92%
Similar papers in this journal
- Molecular correlates and therapeutic targets in T cell-inflamed versus non-T cell-inflamed tumors across cancer types 94%
- Somatic mutational profiles and germline polygenic risk scores in human cancer 93%
- Burden of tumor mutations, neoepitopes, and other variants are dubious predictors of cancer immunotherapy response and overall survival 92%
Similar papers in this journal
- Integrated immunogenomic analyses of high-grade serous ovarian cancer reveal vulnerability to combination immunotherapy 94%
- Tumor reactivity assessment using clonal expression (TRACE) reveals tumor reactive CD8+ T cell heterogeneity across solid tumors 92%
- Machine learning-based single cell and integrative analysis reveals that baseline mDC predisposition predicts protective Hepatitis B vaccine response 92%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.