Accurate sequence-to-affinity models for SH2 domains from multi-round peptide binding assays coupled with free-energy regression
Gagoski, D.; Rube, H. T.; Rastogi, C.; Melo, L. A. N.; Li, X.; Voleti, R.; Shah, N. H.; Bussemaker, H. J.
Show abstract
Short linear peptide motifs play important roles in cell signaling. They can act as modification sites for enzymes and as recognition sites for peptide binding domains. SH2 domains bind specifically to tyrosine-phosphorylated proteins, with the affinity of the interaction depending strongly on the flanking sequence. Quantifying this sequence specificity is critical for deciphering phosphotyrosine-dependent signaling networks. In recent years, protein display technologies and deep sequencing have allowed researchers to profile SH2 domain binding across thousands of candidate ligands. Here, we present a concerted experimental and computational strategy that improves the predictive power of SH2 specificity profiling. Through multi-round affinity selection and deep sequencing with large randomized phosphopeptide libraries, we produce suitable data to train an additive binding free energy model that covers the full theoretical ligand sequence space. Our models can be used to predict signaling network connectivity and the impact of missense variants in phosphoproteins on SH2 binding.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Selection, biophysical and structural analysis of synthetic nanobodies that effectively neutralize SARS-CoV-2 96%
- Identification of motif-based interactions between SARS-CoV-2 protein domains and human peptide ligands pinpoint antiviral targets. 96%
- Allosteric Regulation of the EphA2 Receptor Intracellular Region by Serine/Threonine Kinases 96%
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.