Ultra high diversity factorizable libraries for efficient therapeutic discovery
Dai, Z.; Saksena, S. D.; Horny, G.; Banholzer, C.; Ewert, S.; Gifford, D. K.
Show abstract
The successful discovery of novel biological therapeutics by selection requires highly diverse libraries of candidate sequences that contain a high proportion of desirable candidates. Here we propose the use of computationally designed factorizable libraries made of concatenated segment libraries as a method of creating large libraries that meet an objective function at low cost. We show that factorizable libraries can be designed efficiently by representing objective functions that describe sequence optimality as an inner product of feature vectors, which we use to design an optimization method we call Stochastically Annealed Product Spaces (SAPS). We then use this approach to design diverse and efficient libraries of antibody CDR-H3 sequences with various optimized characteristics.
Matching journals
The top 2 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- FlowDesign: Improved Design of Antibody CDRs Through Flow Matching and Better Prior Distributions 94%
- Sequence-based prediction of protein-protein interactions: a structure-aware interpretable deep learning model 94%
- Meta Learning Improves Robustness and Performance in Machine Learning-Guided Protein Engineering 94%
Similar papers in this journal
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.