Mathematical Characterization of Private and Public Immune Repertoire Sequences
Boettcher, L.; Wald, S.; Chou, T.
Show abstract
Diverse T and B cell repertoires play an important role in mounting effective immune responses against a wide range of pathogens and malignant cells. The number of unique T and B cell clones is characterized by T and B cell receptors (TCRs and BCRs), respectively. Although receptor sequences are generated probabilistically by recombination processes, clinical studies found a high degree of sharing of TCRs and BCRs among different individuals. In this work, we formulate a mathematical and statistical framework to quantify receptor distributions. We define information-theoretic metrics for comparing the frequency of sampled sequences observed across different individuals. Using synthetic and empirical TCR amino acid sequence data, we perform simulations to compare theoretical predictions of this clonal commonality across individuals with corresponding observations. Thus, we quantify the concept of "publicness" or "privateness" of T cell and B cell clones. Our methods can also be used to study the effect of different sampling protocols on the expected commonality of clones and on the confidence levels of this overlap. We also quantify the information loss associated with grouping together certain receptor sequences, as is done in spectratyping.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Fitness dependence of the fixation-time distribution for evolutionary dynamics on graphs 97%
- Mean-field theory accurately captures the variation of copy number distributions across the mRNA's life cycle 97%
- Lifetime of a structure evolving by cluster aggregation and particle loss, and application to postsynaptic scaffold domains 97%
Similar papers in this journal
Similar papers in this journal
- Maximum Mutational Robustness in Genotype-Phenotype Maps Follows a Self-similar Blancmange-like Curve 97%
- Predicting phenotype transition probabilities via conditional algorithmic probability approximations 97%
- A stochastic epidemiological model to estimate the size of an outbreak at the first case identification 96%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.