Non-sequential alignment of binding sites for fast peptide screening
GUYON, F.; Moroy, G.
Show abstract
MotivationPeptides are molecules involved in many essential biological activities by interacting with proteins in the body. They can also be used as therapeutic molecules, particularly to disrupt protein-protein interactions involved in disease. However, it is difficult to identify potential therapeutic peptides if the structure of the protein-protein complex to be inhibited has not been solved. ResultsThe PepIT program was developed to propose peptides that can interact with a given protein. PepIT is based on a non-sequential alignment algorithm to identify peptide binding sites that share geometrical and physicochemical properties with a surface region of the target protein. PepIT compares the entire surface of the target protein with the peptide binding sites of the Propedia dataset, which contains more than 19,000 high-resolution protein-peptide structures. Once a peptide binding site similar to a portion of the protein surface is found, the peptide bound to the binding site is repositioned on the corresponding portion of the protein surface. Availability and implementationThe PepIT source code is freely available at https://github.com/DSIMB/pepit Supplementary informationSupplementary data are available online.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- Estimating statistical significance of local protein profile-profile alignments 96%
- Binding affinity prediction for protein-ligand complex using deep attention mechanism based on intermolecular interactions 96%
- DISTEMA: distance map-based estimation of single protein model accuracy with attentive 2D convolutional neural network 96%
Similar papers in this journal
- Paying Attention to Attention: High Attention Sites as Indicators of Protein Family and Function in Language Models 95%
- Towards a comprehensive view of the pocketome universe - biological implications and algorithmic challenges. 94%
- Novel, provable algorithms for efficient ensemble-based computational protein design and their application to the redesign of the c-Raf-RBD:KRas protein-protein interface 94%
Similar papers in this journal
- Classification of protein binding ligands using structural dispersion of binding site atoms from principal axes 95%
- Using AlphaFold to predict the impact of single mutations on protein stability and function 95%
- SARS-CoV-2 protein structure and sequence mutations: evolutionary analysis and effects on virus variants SARS-CoV-2 protein structure and sequence mutations: 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.