Deep Directed Evolution of Solid Binding Peptides for Quantitative Big-Data Generation
Yucesoy, D. T.; Rath, S. S.; Rodriguez, J. L.; Francis-Landau, J. T.; Nakano-Baker, O.; Sarikaya, M.
Show abstract
Proteins have evolved over millions of years to mediate and carry-out biological processes efficiently. Directed evolution approaches have been used to genetically engineer proteins with desirable functions such as catalysis, mineralization, and target-specific binding. Next-generation sequencing technology offers the capability to discover a massive combinatorial sequence space that is costly to sample experimentally through traditional approaches. Since the permutation space of protein sequence is virtually infinite, and evolution dynamics are poorly understood, experimental verifications have been limited. Recently, machine-learning approaches have been introduced to guide the evolution process that facilitates a deeper and denser search of the sequence-space. Despite these developments, however, frequently used high-fidelity models depend on massive amounts of properly labeled quality data, which so far has been largely lacking in the literature. Here, we provide a preliminary high-throughput peptide-selection protocol with functional scoring to enhance the quality of the data. Solid binding dodecapeptides have been selected against molybdenum disulfide substrate, a two-dimensional atomically thick semiconductor solid. The survival rate of the phage-clones, upon successively stringent washes, quantifies the binding affinity of the peptides onto the solid material. The method suggested here provides a fast generation of preliminary data-pool with [~]2 million unique peptides with 12 amino-acids per sequence by avoiding amplification. Our results demonstrate the importance of data-cleaning and proper conditioning of massive datasets in guiding experiments iteratively. The established extensive groundwork here provides unique opportunities to further iterate and modify the technique to suit a wide variety of needs and generate various peptide and protein datasets. Prospective statistical models developed on the datasets to efficiently explore the sequence-function space will guide towards the intelligent design of proteins and peptides through deep directed evolution. Technological applications of the future based on the peptide-single layer solid based bio/nano soft interfaces, such as biosensors, bioelectronics, and logic devices, is expected to benefit from the solid binding peptide dataset alone. Furthermore, protocols described herein will also benefit efforts in medical applications, such as vaccine development, that could significantly accelerate a global response to future pandemics.
Matching journals
The top 11 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Engineering a virus-like particle to display peptide insertions using an apparent fitness landscape 95%
- Exploring the Extreme Acid Tolerance of a Dynamic Protein Nanocage 92%
- Applying computational protein design to engineer affibodies for affinity-controlled delivery of vascular endothelial growth factor and platelet-derived growth factor 91%
Similar papers in this journal
Similar papers in this journal
- Learning Peptide Recognition Rules for a Low-Specificity Protein 94%
- Phage display assisted discovery of a pH-dependent anti-alpha-cobratoxin antibody from a natural variable domain library 93%
- Screening de novo designed protein binders in unpurified lysate using flow induced dispersion analysis 92%
Similar papers in this journal
- Systematic profiling of peptide substrate specificity in N-terminal processing by methionine aminopeptidase using mRNA display and an unnatural methionine analogue 95%
- Peptide-antibody Fusions Engineered by Phage Display Exhibit Ultrapotent and Broad Neutralization of SARS-CoV-2 Variants 93%
- A Novel Regioselective Approach to Cyclize Phage-Displayed Peptides in Combination with Epitope-Directed Selection to Identify a Potent Neutralizing Macrocyclic Peptide for SARS-CoV-2 92%
Similar papers in this journal
- Assembly mechanism of the AIM2 inflammasome sensorrevealed by single-molecule analysis 94%
- Machine Learning Optimization of Candidate Antibodies Yields Highly Diverse Sub-nanomolar Affinity Antibody Libraries 93%
- Notch engagement by Jag1 nanoscale clusters indicates a force-independent mode of activation 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.