Learning Human T Cell Behaviors through Generative AI Embeddings of T Cell Receptors
Chen, D. G.; Su, Y.; Heath, J. R.
Show abstract
T cells interact with the world through T cell receptors (TCRs). The extent to which TCRs determine T cell behavior has not been comprehensively characterized. Our Tarpon model leverages advances in generative artificial intelligence to synthesize large-scale (>1M sequences) TCR atlases across human development and diseases into actionable insights. Tarpon creates: 1) bespoke sampling functions generating realistic Ag-specific TCRs, 2) embeddings revealing CD4+ and CD8+ single-positive TCR repertoires as distinct with divergent physiochemical properties, and 3) cross-dataset mappings of T cell states that validate fetal CD4+ versus CD8+ TCR differences in adults and find fetal type I innate T cells to map to MAIT and KIR+ adult CD8+ T cells which we verify via whole transcriptome analysis. Tarpon is a resource as a reference of TCRs across human physiological states and as a computational framework to create interpretable TCR embeddings, via physicochemical associations, that have broad implications for the field.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- APMAT analysis reveals the association between CD8 T cell receptors, cognate antigen, and T cell phenotype and persistence 98%
- Projecting single-cell transcriptomics data onto a reference T cell atlas to interpret immune responses 97%
- Deep learning predictions of TCR-epitope interactions reveal epitope-specific chains in dual alpha T cells 97%
Similar papers in this journal
- Gene regulatory network inference from CRISPR perturbations in primary CD4+ T cells elucidates the genomic basis of immune disease 97%
- Impact of disease-associated chromatin accessibility QTLs across immune cell types and contexts 97%
- Interpretable deep learning reveals the sequence rules of Hippo signaling 96%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.