Design of linear and cyclic peptide binders of different lengths from protein sequence information
Li, Q.; Vlachos, E. N.; Bryant, P.
10.1101/2024.06.20.599739 bioRxivShow abstract
Structure prediction technology has revolutionised the field of protein design, but key questions such as how to design new functions remain. Many proteins exert their functions through interactions with other proteins, and a significant challenge is designing these interactions effectively. While most efforts have focused on larger, more stable proteins, shorter peptides offer advantages such as lower manufacturing costs, reduced steric hindrance, and the ability to traverse cell membranes when cyclized. However, less structural data is available for peptides and their flexibility makes them harder to design. Here, we present a method to design both novel linear and cyclic peptide binders of varying lengths based solely on a protein target sequence. Our approach does not specify a binding site or the length of the binder, making the procedure completely blind. We demonstrate that linear and cyclic peptide binders of different lengths can be designed with nM affinity in a single shot, and adversarial designs can be avoided through orthogonal in silico evaluation, tripling the success rate. Our protocol, EvoBind2 is freely available https://github.com/patrickbryant1/EvoBind.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- AbDesign: Database of point mutants of antibodies with associated structures reveals poor generalization of binding predictions from machine learning models. 95%
- Towards generalizable prediction of antibody thermostability using machine learning on sequence and structure features 94%
- AlphaBind, a Domain-Specific Model to Predict and Optimize Antibody-Antigen Binding Affinity 94%
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.